Intelligence Index
AA · Chat / text · Combined
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #259 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 12
- Percentile
- 53.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `intelligenceIndex`.
53.5% percentile inside its fair comparison set12Raw benchmark value
AA-Omniscience accuracy
AA · Chat / text · Objective
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #310 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 15.7%
- Percentile
- 32%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `omniscienceAccuracy`.
32% percentile inside its fair comparison set15.7%Raw benchmark value
AA-Omniscience non-hallucination
AA · Chat / text · Objective
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #75 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 55.1%
- Percentile
- 83.7%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `omniscienceNonHallucination`.
83.7% percentile inside its fair comparison set55.1%Raw benchmark value
IFBench
AA · Chat / text · Objective
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #188 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 40.5%
- Percentile
- 46.3%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `ifbench`.
46.3% percentile inside its fair comparison set40.5%Raw benchmark value
Input price
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #255 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- $1.5 /1M input tokens
- Percentile
- 28.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `price1mInputTokens`.
28.5% percentile inside its fair comparison set$1.5 /1M input tokensRaw benchmark value
Output price
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #262 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- $7.5 /1M output tokens
- Percentile
- 26.3%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `price1mOutputTokens`.
26.3% percentile inside its fair comparison set$7.5 /1M output tokensRaw benchmark value
Output Speed
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #212 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 53.7 tokens/s
- Percentile
- 19.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `medianOutputTokensPerSecond`.
19.5% percentile inside its fair comparison set53.7 tokens/sRaw benchmark value
Time to first token
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #144 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 3.05s
- Percentile
- 45.8%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `medianTimeToFirstTokenSeconds`.
45.8% percentile inside its fair comparison set3.05sRaw benchmark value
Time to first answer token
AA · Chat / text · Speed / cost
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #62 · Source label: Qwen3 Coder 480B A35B Instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Artificial Analysis
- Raw value
- 3.05s
- Percentile
- 76.7%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Artificial Analysis public leaderboard field `medianTimeToFirstAnswerTokenSeconds`.
76.7% percentile inside its fair comparison set3.05sRaw benchmark value
Text Arena
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #159 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,387
- Percentile
- 58%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: overall. Source rank: #178. Votes: 25006. Organization: alibaba. License: Apache 2.0.
58% percentile inside its fair comparison set1,387Raw benchmark valueCI 1,382 - 1,392
Text Arena · Creative Writing
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #145 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,363
- Percentile
- 61.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: creative_writing. Source rank: #164. Votes: 3193. Organization: alibaba. License: Apache 2.0.
61.5% percentile inside its fair comparison set1,363Raw benchmark valueCI 1,352 - 1,373
Text Arena · English
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #164 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,396
- Percentile
- 56.6%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: english. Source rank: #183. Votes: 10891. Organization: alibaba. License: Apache 2.0.
56.6% percentile inside its fair comparison set1,396Raw benchmark valueCI 1,390 - 1,403
Text Arena · Exclude Ties
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #159 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,364
- Percentile
- 58%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: exclude_ties. Source rank: #178. Votes: 17860. Organization: alibaba. License: Apache 2.0.
58% percentile inside its fair comparison set1,364Raw benchmark valueCI 1,357 - 1,370
Text Arena · Hard Prompts
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #152 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,414
- Percentile
- 59.8%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: hard_prompts. Source rank: #170. Votes: 11452. Organization: alibaba. License: Apache 2.0.
59.8% percentile inside its fair comparison set1,414Raw benchmark valueCI 1,407 - 1,420
Text Arena · Hard Prompts English
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #146 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,425
- Percentile
- 61.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: hard_prompts_english. Source rank: #162. Votes: 5100. Organization: alibaba. License: Apache 2.0.
61.2% percentile inside its fair comparison set1,425Raw benchmark valueCI 1,416 - 1,434
Text Arena · Instruction Following
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #148 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,384
- Percentile
- 60.9%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: instruction_following. Source rank: #167. Votes: 6145. Organization: alibaba. License: Apache 2.0.
60.9% percentile inside its fair comparison set1,384Raw benchmark valueCI 1,376 - 1,391
Text Arena · Longer Query
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #138 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,408
- Percentile
- 61.4%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: longer_query. Source rank: #157. Votes: 4887. Organization: alibaba. License: Apache 2.0.
61.4% percentile inside its fair comparison set1,408Raw benchmark valueCI 1,399 - 1,417
Text Arena · Multi Turn
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #146 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,397
- Percentile
- 61.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: multi_turn. Source rank: #165. Votes: 4251. Organization: alibaba. License: Apache 2.0.
61.2% percentile inside its fair comparison set1,397Raw benchmark valueCI 1,388 - 1,407
Text Arena · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #176 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,356
- Percentile
- 53.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: overall. Source rank: #197. Votes: 25006. Organization: alibaba. License: Apache 2.0.
53.5% percentile inside its fair comparison set1,356Raw benchmark valueCI 1,351 - 1,361
Text Arena · Creative Writing · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #161 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,333
- Percentile
- 57.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: creative_writing. Source rank: #181. Votes: 3193. Organization: alibaba. License: Apache 2.0.
57.2% percentile inside its fair comparison set1,333Raw benchmark valueCI 1,323 - 1,344
Text Arena · English · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #190 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,359
- Percentile
- 49.7%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: english. Source rank: #211. Votes: 10891. Organization: alibaba. License: Apache 2.0.
49.7% percentile inside its fair comparison set1,359Raw benchmark valueCI 1,353 - 1,366
Text Arena · Exclude Ties · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #176 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,318
- Percentile
- 53.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: exclude_ties. Source rank: #197. Votes: 17860. Organization: alibaba. License: Apache 2.0.
53.5% percentile inside its fair comparison set1,318Raw benchmark valueCI 1,311 - 1,325
Text Arena · Hard Prompts · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #162 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,372
- Percentile
- 57.2%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: hard_prompts. Source rank: #182. Votes: 11452. Organization: alibaba. License: Apache 2.0.
57.2% percentile inside its fair comparison set1,372Raw benchmark valueCI 1,366 - 1,378
Text Arena · Hard Prompts English · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #164 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,381
- Percentile
- 56.4%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: hard_prompts_english. Source rank: #183. Votes: 5100. Organization: alibaba. License: Apache 2.0.
56.4% percentile inside its fair comparison set1,381Raw benchmark valueCI 1,372 - 1,389
Text Arena · Instruction Following · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #159 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,356
- Percentile
- 58%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: instruction_following. Source rank: #178. Votes: 6145. Organization: alibaba. License: Apache 2.0.
58% percentile inside its fair comparison set1,356Raw benchmark valueCI 1,348 - 1,364
Text Arena · Longer Query · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #152 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,378
- Percentile
- 57.5%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: longer_query. Source rank: #171. Votes: 4887. Organization: alibaba. License: Apache 2.0.
57.5% percentile inside its fair comparison set1,378Raw benchmark valueCI 1,370 - 1,387
Text Arena · Multi Turn · No Style Control
AR · Chat / text · Human
It tests whether the model is actually useful in normal conversational turns, not just on narrow correctness tasks.
Rank #164 · Source label: qwen3-coder-480b-a35b-instruct
verified runtimeexact alias
Raw row drilldownsource row, percentile, last updated, eligibility
- Source
- Arena
- Raw value
- 1,365
- Percentile
- 56.4%
- Last updated
- recent
- Eligibility
- headline eligible
Parsed from Arena leaderboard dataset row `qwen3-coder-480b-a35b-instruct`. Category: multi_turn. Source rank: #182. Votes: 4251. Organization: alibaba. License: Apache 2.0.
56.4% percentile inside its fair comparison set1,365Raw benchmark valueCI 1,356 - 1,375