Qwen3 VL 235B A22B (Reasoning)
Qwen
- Chat / text
- Reasoning / math / science
Qwen
0 shared benchmarks are still too close to call, so the win stays conditional. This compare uses all public sources, with provider-official evidence labeled separately.
Unknown
| Intelligence Index AA · index Text · Chat / text | 1357.8% exact aliasverified runtime Row details
| 23n/a exact aliasverified runtimeContext only Row details
| n/a |
| GPQA AA · % Text · Reasoning / math / science | 77.2%68.6% exact aliasverified runtime Row details
| 85.7%n/a exact aliasverified runtimeContext only Row details
| n/a |
| Humanity's Last Exam AA · % Text · Reasoning / math / science | 11.9%61.8% exact aliasverified runtime Row details
| 29.6%n/a exact aliasverified runtimeContext only Row details
| n/a |
| CritPt AA · % Text · Reasoning / math / science | 0%50.1% exact aliasverified runtime Row details
| 8.9%n/a exact aliasverified runtimeContext only Row details
| n/a |
| AA-Omniscience accuracy AA · % Text · Chat / text | 20.9%53.2% exact aliasverified runtime Row details
| 18.6%n/a exact aliasverified runtimeContext only Row details
| n/a |
| AA-Omniscience non-hallucination AA · % Text · Chat / text | 14.8%41.3% exact aliasverified runtime Row details
| 67%n/a exact aliasverified runtimeContext only Row details
| n/a |
| Long Context Reasoning AA · % Document · Long context | 63.7%60.4% exact aliasverified runtime Row details
| 70%n/a exact aliasverified runtimeContext only Row details
| n/a |