Quality
Known Good Index and the capability indices, side by side
Each index is min-max normalised across tracked models and scaled 0–100 from published evaluations. We do not run these evaluations.
| Metric | Anthropic: Claude Opus 4.6 | Google: Gemini 2.5 Pro |
|---|---|---|
| Known Good Index | 88 | 64 |
| Coding Index | 89 | 42 |
| Mathematics Index | 93 | 84 |
| Science Index | 95 | 89 |
| Reasoning Index | 84 | — |
| Agentic Index | 89 | 42 |
| Openness Index | 5 | 5 |
Quality Breakdown
Every evaluation both models report a published score for
Bar = Anthropic: Claude Opus 4.6 · marker = Google: Gemini 2.5 Pro · scores ingested from Epoch AI (CC-BY 4.0)
Right-hand figures read Anthropic: Claude Opus 4.6 / Google: Gemini 2.5 Pro. Anthropic: Claude Opus 4.6 leads on 4 of 4, Google: Gemini 2.5 Pro on 0.
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena. Higher is better.
| Metric | Anthropic: Claude Opus 4.6 | Google: Gemini 2.5 Pro |
|---|---|---|
| Text Arena | 1,497 | 1,457 |
| Text Arena (style controlled) | 1,498 | 1,446 |
Price & Cost
Published API pricing by token type
USD per 1M tokens. Blended is a 3:1 input:output mix. Live from OpenRouter, cross-checked against models.dev. Lower is better.
| Metric | Anthropic: Claude Opus 4.6 | Google: Gemini 2.5 Pro |
|---|---|---|
| Input $/1M | $5.00 | $1.25 |
| Output $/1M | $25.00 | $10.00 |
| Blended $/1M | $10.00 | $3.44 |
| Cache read $/1M | $0.500 | $0.125 |
Context Window
Maximum input accepted and maximum output emitted
Provider-declared limits, from metadata via models.dev and OpenRouter. Higher is better.
| Metric | Anthropic: Claude Opus 4.6 | Google: Gemini 2.5 Pro |
|---|---|---|
| Context window | 1.00M | 1.05M |
| Max output tokens | 128k | 66k |
Speed & Latency
Median throughput and time to first token across hosting providers
Medians across all tracked providers, derived from OpenRouter endpoint telemetry.
| Metric | Anthropic: Claude Opus 4.6 | Google: Gemini 2.5 Pro |
|---|---|---|
| Output tokens/s | 37 | 96 |
| Time to first token | 2.1s | 5.4s |
Providers
Which hosts serve each model, and at what price and speed
| Provider | Anthropic: Claude Opus 4.6 blended | Anthropic: Claude Opus 4.6 tok/s | Google: Gemini 2.5 Pro blended | Google: Gemini 2.5 Pro tok/s |
|---|---|---|---|---|
| Google · first party | $10.00 | 40 | $1.72 | 99 |
| Anthropic · first party | $10.00 | 37 | — | — |
| Amazon Bedrock · first party | $10.00 | 33 | — | — |
| Azure · first party | $10.00 | 30 | — | — |
| Google AI Studio | — | — | $1.72 | 118 |
Per-host pricing and telemetry from OpenRouter. A dash means that host does not serve that model.