| Time to first answer token AA · s Text · Chat / text | 57.61s14.1%exact aliasverified runtime Row details- Raw value
- 57.61s
- Percentile
- 14.1%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 25.34s40.1%exact aliasverified runtime Row details- Raw value
- 25.34s
- Percentile
- 40.1%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Adaptive Reasoning, Max Effort)
Source row | |
| MMMU-Pro AA · % Vision · Vision understanding | 84.9%94.2%exact aliasverified runtime Row details- Raw value
- 84.9%
- Percentile
- 94.2%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 76.4%71.6%exact aliasverified runtime Row details- Raw value
- 76.4%
- Percentile
- 71.6%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| CritPt AA · % Text · Reasoning / math / science | 30%97.1%exact aliasverified runtime Row details- Raw value
- 30%
- Percentile
- 97.1%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 5.1%77.9%exact aliasverified runtime Row details- Raw value
- 5.1%
- Percentile
- 77.9%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| Input price AA · /1M input tokens Text · Chat / text | $2 /1M input tokens26.8%exact aliasverified runtime Row details- Raw value
- $2 /1M input tokens
- Percentile
- 26.8%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | $5 /1M input tokens9.3%exact aliasverified runtime Row details- Raw value
- $5 /1M input tokens
- Percentile
- 9.3%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Adaptive Reasoning, Max Effort)
Source row | |
| Output Speed AA · tokens/s Text · Chat / text | 58.2 tokens/s26.7%exact aliasverified runtime Row details- Raw value
- 58.2 tokens/s
- Percentile
- 26.7%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 43.2 tokens/s9.5%exact aliasverified runtime Row details- Raw value
- 43.2 tokens/s
- Percentile
- 9.5%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| Long Context Reasoning AA · % Document · Long context | 82.3%94.3%exact aliasverified runtime Row details- Raw value
- 82.3%
- Percentile
- 94.3%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 75.7%77.5%exact aliasverified runtime Row details- Raw value
- 75.7%
- Percentile
- 77.5%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| Output price AA · /1M output tokens Text · Chat / text | $10 /1M output tokens23.7%exact aliasverified runtime Row details- Raw value
- $10 /1M output tokens
- Percentile
- 23.7%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | $25 /1M output tokens9%exact aliasverified runtime Row details- Raw value
- $25 /1M output tokens
- Percentile
- 9%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Adaptive Reasoning, Max Effort)
Source row | |
| Humanity's Last Exam AA · % Text · Reasoning / math / science | 51.4%96.2%exact aliasverified runtime Row details- Raw value
- 51.4%
- Percentile
- 96.2%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 33.3%82.6%exact aliasverified runtime Row details- Raw value
- 33.3%
- Percentile
- 82.6%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| Intelligence Index AA · index Text · Chat / text | 5097.1%exact aliasverified runtime Row details- Raw value
- 50
- Percentile
- 97.1%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 3185.4%exact aliasverified runtime Row details- Raw value
- 31
- Percentile
- 85.4%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| AA-Omniscience accuracy AA · % Text · Chat / text | 60.8%96.5%exact aliasverified runtime Row details- Raw value
- 60.8%
- Percentile
- 96.5%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 44.7%85.8%exact aliasverified runtime Row details- Raw value
- 44.7%
- Percentile
- 85.8%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |
| Time to first token AA · s Text · Chat / text | 57.61s9.8%exact aliasverified runtime Row details- Raw value
- 57.61s
- Percentile
- 9.8%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 25.34s17%exact aliasverified runtime Row details- Raw value
- 25.34s
- Percentile
- 17%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Adaptive Reasoning, Max Effort)
Source row | |
| AA-Omniscience non-hallucination AA · % Text · Chat / text | 50.6%81.2%exact aliasverified runtime Row details- Raw value
- 50.6%
- Percentile
- 81.2%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- GPT-6.1 Sol (High)
Source row | 45.9%76.8%exact aliasverified runtime Row details- Raw value
- 45.9%
- Percentile
- 76.8%
- Last updated
- recent
- Eligibility
- headline eligible
- Identity
- provider alias (0.94)
- Source label
- Claude Opus 4.7 (Non-reasoning, High Effort)
Source row | |