Gemini 2.5 Pro Preview (Mar' 25)
- Reasoning / math / science
0 shared benchmarks are still too close to call, so the win stays conditional. This compare uses all public sources, with provider-official evidence labeled separately.
Unknown
| Intelligence Index AA · index Text · Chat / text | 23n/a exact aliasverified runtimeContext only Row details
| 40n/a exact aliasverified runtimeContext only Row details
| n/a |
| Coding Index AA · index Code · Coding | 47n/a exact aliasverified runtimeContext only Row details
| 59n/a exact aliasverified runtimeContext only Row details
| n/a |
| GPQA AA · % Text · Reasoning / math / science | 83.6%n/a exact aliasverified runtimeContext only Row details
| 87.6%n/a exact aliasverified runtimeContext only Row details
| n/a |
| Humanity's Last Exam AA · % Text · Reasoning / math / science | 18%n/a exact aliasverified runtimeContext only Row details
| 33.6%n/a exact aliasverified runtimeContext only Row details
| n/a |
| SciCode AA · % Code · Coding | 39.5%n/a exact aliasverified runtimeContext only Row details
| 42.2%n/a exact aliasverified runtimeContext only Row details
| n/a |