| Intelligence Index AA · index Text · Chat / text | 26n/aexact aliasverified runtimeContext only Row details- Raw value
- 26
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 35n/aexact aliasverified runtimeContext only Row details- Raw value
- 35
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| Openness Index AA · index Text · Chat / text | 17n/aexact aliasverified runtimeContext only Row details- Raw value
- 17
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 50n/aexact aliasverified runtimeContext only Row details- Raw value
- 50
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| GPQA AA · % Text · Reasoning / math / science | 77.6%n/aexact aliasverified runtimeContext only Row details- Raw value
- 77.6%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 85.7%n/aexact aliasverified runtimeContext only Row details- Raw value
- 85.7%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| Humanity's Last Exam AA · % Text · Reasoning / math / science | 12.7%n/aexact aliasverified runtimeContext only Row details- Raw value
- 12.7%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 29.6%n/aexact aliasverified runtimeContext only Row details- Raw value
- 29.6%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| CritPt AA · % Text · Reasoning / math / science | 0%n/aexact aliasverified runtimeContext only Row details- Raw value
- 0%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 8.9%n/aexact aliasverified runtimeContext only Row details- Raw value
- 8.9%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| SciCode AA · % Code · Coding | 38.7%n/aexact aliasverified runtimeContext only Row details- Raw value
- 38.7%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 38.5%n/aexact aliasverified runtimeContext only Row details- Raw value
- 38.5%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| AA-Omniscience accuracy AA · % Text · Chat / text | 27.4%n/aexact aliasverified runtimeContext only Row details- Raw value
- 27.4%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 18.6%n/aexact aliasverified runtimeContext only Row details- Raw value
- 18.6%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| AA-Omniscience non-hallucination AA · % Text · Chat / text | 5.3%n/aexact aliasverified runtimeContext only Row details- Raw value
- 5.3%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking
Source row | 67%n/aexact aliasverified runtimeContext only Row details- Raw value
- 67%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |
| Long Context Reasoning AA · % Document · Long context | 60.3%n/aexact aliasverified runtimeContext only Row details- Raw value
- 60.3%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- preview_model
- Identity
- provider alias (0.94)
- Source label
- Qwen3 Max Thinking (Preview)
Source row | 66%n/aexact aliasverified runtimeContext only Row details- Raw value
- 66%
- Percentile
- n/a
- Last updated
- recent
- Eligibility
- benchmark_derived_model
- Identity
- provider alias (0.94)
- Source label
- A.X-K2
Source row | |