| Intelligence Index AA · index Text · Chat / text | 4091.4%exact aliasverified runtime Row details- Raw value
- 40
- Percentile
- 91.4%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 23n/aexact aliasverified runtimeContext only Row details- Raw value
- 23
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| GPQA AA · % Text · Reasoning / math / science | 92.7%94.7%exact aliasverified runtime Row details- Raw value
- 92.7%
- Percentile
- 94.7%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 85.7%n/aexact aliasverified runtimeContext only Row details- Raw value
- 85.7%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| Humanity's Last Exam AA · % Text · Reasoning / math / science | 43%91.1%exact aliasverified runtime Row details- Raw value
- 43%
- Percentile
- 91.1%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 29.6%n/aexact aliasverified runtimeContext only Row details- Raw value
- 29.6%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| CritPt AA · % Text · Reasoning / math / science | 20%90.8%exact aliasverified runtime Row details- Raw value
- 20%
- Percentile
- 90.8%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 8.9%n/aexact aliasverified runtimeContext only Row details- Raw value
- 8.9%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| SciCode AA · % Code · Coding | 53.2%70.2%exact aliasverified runtime Row details- Raw value
- 53.2%
- Percentile
- 70.2%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 41%n/aexact aliasverified runtimeContext only Row details- Raw value
- 41%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| AA-Omniscience accuracy AA · % Text · Chat / text | 31.9%75.5%exact aliasverified runtime Row details- Raw value
- 31.9%
- Percentile
- 75.5%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 18.6%n/aexact aliasverified runtimeContext only Row details- Raw value
- 18.6%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| AA-Omniscience non-hallucination AA · % Text · Chat / text | 58.3%85.4%exact aliasverified runtime Row details- Raw value
- 58.3%
- Percentile
- 85.4%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 67%n/aexact aliasverified runtimeContext only Row details- Raw value
- 67%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| Long Context Reasoning AA · % Document · Long context | 78.3%82.6%exact aliasverified runtime Row details- Raw value
- 78.3%
- Percentile
- 82.6%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Qwen3.8 Max
Source row | 70%n/aexact aliasverified runtimeContext only Row details- Raw value
- 70%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |