benchmarks
Top Models by Benchmark Score (2026)
Ranked by blended benchmark data from Aider Polyglot and Artificial Analysis Intelligence Index. Available models only - any under access restrictions are excluded.
| # | Model | Blended |
|---|---|---|
| 1 | Anthropic: Claude Opus 5.5 (batch) | 100.0 |
| 2 | Anthropic: Claude Opus 5.5 | 100.0 |
| 3 | Google: Gemini 2.5 Pro | 94.2 |
| 4 | Google: Gemini 2.5 Pro (batch) | 94.2 |
| 5 | Anthropic: Claude Fable 5.1 (batch) | 92.0 |