Intelligence Index
AA · Chat / text · Combined
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #132 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 24
- Percentile
- 72%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `intelligenceIndex`.
72% percentile inside its fair comparison set24Raw benchmark value
AA-Omniscience accuracy
AA · Chat / text · Objective
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #262 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 14.4%
- Percentile
- 29%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `omniscienceAccuracy`.
29% percentile inside its fair comparison set14.4%Raw benchmark value
AA-Omniscience non-hallucination
AA · Chat / text · Objective
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #10 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 74.3%
- Percentile
- 97.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `omniscienceNonHallucination`.
97.5% percentile inside its fair comparison set74.3%Raw benchmark value
IFBench
AA · Chat / text · Objective
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #167 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 42%
- Percentile
- 52.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `ifbench`.
52.2% percentile inside its fair comparison set42%Raw benchmark value
Blended price
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #219 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- $2 /1M tokens
- Percentile
- 28.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `price1mBlended0To3To1`.
28.2% percentile inside its fair comparison set$2 /1M tokensRaw benchmark value
Input price
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #210 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- $1 /1M input tokens
- Percentile
- 32.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `price1mInputTokens`.
32.2% percentile inside its fair comparison set$1 /1M input tokensRaw benchmark value
Output price
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #226 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- $5 /1M output tokens
- Percentile
- 25.9%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `price1mOutputTokens`.
25.9% percentile inside its fair comparison set$5 /1M output tokensRaw benchmark value
Output Speed
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #111 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 90.1 tokens/s
- Percentile
- 50.9%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `medianOutputTokensPerSecond`.
50.9% percentile inside its fair comparison set90.1 tokens/sRaw benchmark value
Time to first token
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #39 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 1.04s
- Percentile
- 83%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `medianTimeToFirstTokenSeconds`.
83% percentile inside its fair comparison set1.04sRaw benchmark value
Time to first answer token
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #27 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 1.04s
- Percentile
- 88.4%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `medianTimeToFirstAnswerTokenSeconds`.
88.4% percentile inside its fair comparison set1.04sRaw benchmark value
Openness Index
AA · Chat / text · Combined
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #240 · Source label: Claude 4.5 Haiku (Non-reasoning)
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 11
- Percentile
- 11.3%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `opennessBreakdown.opennessIndex`.
11.3% percentile inside its fair comparison set11Raw benchmark value
Text Arena
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #86 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,411
- Percentile
- 73.8%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: overall. Source rank: #106. Votes: 91153. Organization: anthropic. License: Proprietary.
73.8% percentile inside its fair comparison set1,411Raw benchmark valueCI 1,408 - 1,414
Text Arena · Creative Writing
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #78 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,386
- Percentile
- 76.2%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: creative_writing. Source rank: #96. Votes: 13641. Organization: anthropic. License: Proprietary.
76.2% percentile inside its fair comparison set1,386Raw benchmark valueCI 1,381 - 1,392
Text Arena · English
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #71 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,432
- Percentile
- 78.5%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: english. Source rank: #89. Votes: 43281. Organization: anthropic. License: Proprietary.
78.5% percentile inside its fair comparison set1,432Raw benchmark valueCI 1,428 - 1,435
Text Arena · Exclude Ties
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #86 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,398
- Percentile
- 73.8%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: exclude_ties. Source rank: #106. Votes: 67483. Organization: anthropic. License: Proprietary.
73.8% percentile inside its fair comparison set1,398Raw benchmark valueCI 1,394 - 1,402
Text Arena · Hard Prompts
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #73 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,439
- Percentile
- 77.8%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: hard_prompts. Source rank: #91. Votes: 53244. Organization: anthropic. License: Proprietary.
77.8% percentile inside its fair comparison set1,439Raw benchmark valueCI 1,435 - 1,443
Text Arena · Hard Prompts English
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #63 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,455
- Percentile
- 80.9%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: hard_prompts_english. Source rank: #79. Votes: 26560. Organization: anthropic. License: Proprietary.
80.9% percentile inside its fair comparison set1,455Raw benchmark valueCI 1,450 - 1,459
Text Arena · Instruction Following
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #69 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,414
- Percentile
- 79.1%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: instruction_following. Source rank: #86. Votes: 26995. Organization: anthropic. License: Proprietary.
79.1% percentile inside its fair comparison set1,414Raw benchmark valueCI 1,410 - 1,419
Text Arena · Longer Query
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #61 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,437
- Percentile
- 80.3%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: longer_query. Source rank: #76. Votes: 29096. Organization: anthropic. License: Proprietary.
80.3% percentile inside its fair comparison set1,437Raw benchmark valueCI 1,432 - 1,441
Text Arena · Multi Turn
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #68 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,426
- Percentile
- 79.3%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: multi_turn. Source rank: #86. Votes: 16101. Organization: anthropic. License: Proprietary.
79.3% percentile inside its fair comparison set1,426Raw benchmark valueCI 1,420 - 1,431
Text Arena · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #101 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,392
- Percentile
- 69.2%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: overall. Source rank: #122. Votes: 91153. Organization: anthropic. License: Proprietary.
69.2% percentile inside its fair comparison set1,392Raw benchmark valueCI 1,389 - 1,395
Text Arena · Creative Writing · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #82 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,369
- Percentile
- 74.9%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: creative_writing. Source rank: #102. Votes: 13641. Organization: anthropic. License: Proprietary.
74.9% percentile inside its fair comparison set1,369Raw benchmark valueCI 1,363 - 1,375
Text Arena · English · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #92 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,413
- Percentile
- 72%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: english. Source rank: #108. Votes: 43281. Organization: anthropic. License: Proprietary.
72% percentile inside its fair comparison set1,413Raw benchmark valueCI 1,409 - 1,417
Text Arena · Exclude Ties · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #101 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,370
- Percentile
- 69.2%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: exclude_ties. Source rank: #122. Votes: 67483. Organization: anthropic. License: Proprietary.
69.2% percentile inside its fair comparison set1,370Raw benchmark valueCI 1,367 - 1,374
Text Arena · Hard Prompts · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #80 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,417
- Percentile
- 75.7%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: hard_prompts. Source rank: #97. Votes: 53244. Organization: anthropic. License: Proprietary.
75.7% percentile inside its fair comparison set1,417Raw benchmark valueCI 1,413 - 1,420
Text Arena · Hard Prompts English · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #65 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,434
- Percentile
- 80.2%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: hard_prompts_english. Source rank: #80. Votes: 26560. Organization: anthropic. License: Proprietary.
80.2% percentile inside its fair comparison set1,434Raw benchmark valueCI 1,429 - 1,439
Text Arena · Instruction Following · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #50 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,412
- Percentile
- 84.9%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: instruction_following. Source rank: #63. Votes: 26995. Organization: anthropic. License: Proprietary.
84.9% percentile inside its fair comparison set1,412Raw benchmark valueCI 1,407 - 1,416
Text Arena · Longer Query · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #49 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,427
- Percentile
- 84.2%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: longer_query. Source rank: #61. Votes: 29096. Organization: anthropic. License: Proprietary.
84.2% percentile inside its fair comparison set1,427Raw benchmark valueCI 1,423 - 1,432
Text Arena · Multi Turn · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #81 · Source label: claude-haiku-4-5-20251001
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,408
- Percentile
- 75.2%
- Last updated
- aging
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `claude-haiku-4-5-20251001`. Category: multi_turn. Source rank: #98. Votes: 16101. Organization: anthropic. License: Proprietary.
75.2% percentile inside its fair comparison set1,408Raw benchmark valueCI 1,403 - 1,414