Gemini 2.5 Pro Preview (Mar' 25)
- Reasoning / math / science
0 shared benchmarks are still too close to call, so the win stays conditional. This compare uses all public sources, with provider-official evidence labeled separately.
Unknown
| Intelligence Index AA · index Text · Chat / text | 23n/a exact aliasverified runtimeContext only Row details
| 35n/a exact aliasverified runtimeContext only Row details
| n/a |
| Coding Index AA · index Code · Coding | 47n/a exact aliasverified runtimeContext only Row details
| 39n/a exact aliasverified runtimeContext only Row details
| n/a |
| GPQA AA · % Text · Reasoning / math / science | 83.6%n/a exact aliasverified runtimeContext only Row details
| 85.7%n/a exact aliasverified runtimeContext only Row details
| n/a |
| Humanity's Last Exam AA · % Text · Reasoning / math / science | 18%n/a exact aliasverified runtimeContext only Row details
| 29.6%n/a exact aliasverified runtimeContext only Row details
| n/a |
| SciCode AA · % Code · Coding | 39.5%n/a exact aliasverified runtimeContext only Row details
| 38.5%n/a exact aliasverified runtimeContext only Row details
| n/a |