Quality
Known Good Index and the capability indices, side by side
Each index is min-max normalised across tracked models and scaled 0–100 from published evaluations. We do not run these evaluations.
| Metric | claude-opus-4-20250514 | MoonshotAI: Kimi K2.6 |
|---|---|---|
| Known Good Index | 62 | 93 |
| Mathematics Index | 64 | 96 |
| Science Index | 69 | 95 |
| Agentic Index | 76 | 87 |
| Openness Index | 5 | 60 |
Quality Breakdown
Every evaluation both models report a published score for
Bar = claude-opus-4-20250514 · marker = MoonshotAI: Kimi K2.6 · scores ingested from Epoch AI (CC-BY 4.0)
Right-hand figures read claude-opus-4-20250514 / MoonshotAI: Kimi K2.6. claude-opus-4-20250514 leads on 0 of 3, MoonshotAI: Kimi K2.6 on 3.
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena. Higher is better.
| Metric | claude-opus-4-20250514 | MoonshotAI: Kimi K2.6 |
|---|---|---|
| Text Arena | 1,364 | 1,455 |
| Text Arena (style controlled) | 1,412 | 1,461 |
Price & Cost
Published API pricing by token type
USD per 1M tokens. Blended is a 3:1 input:output mix. Live from OpenRouter, cross-checked against models.dev. Lower is better.
| Metric | claude-opus-4-20250514 | MoonshotAI: Kimi K2.6 |
|---|---|---|
| Input $/1M | — | $0.65 |
| Output $/1M | — | $2.72 |
| Blended $/1M | — | $1.16 |
| Cache read $/1M | — | $0.109 |
Context Window
Maximum input accepted and maximum output emitted
Provider-declared limits, from metadata via models.dev and OpenRouter. Higher is better.
| Metric | claude-opus-4-20250514 | MoonshotAI: Kimi K2.6 |
|---|---|---|
| Context window | — | 262k |
| Max output tokens | — | 262k |
Speed & Latency
Median throughput and time to first token across hosting providers
Medians across all tracked providers, derived from OpenRouter endpoint telemetry.
| Metric | claude-opus-4-20250514 | MoonshotAI: Kimi K2.6 |
|---|---|---|
| Output tokens/s | — | 66 |
| Time to first token | — | 1.1s |
Providers
Which hosts serve each model, and at what price and speed
| Provider | claude-opus-4-20250514 blended | claude-opus-4-20250514 tok/s | MoonshotAI: Kimi K2.6 blended | MoonshotAI: Kimi K2.6 tok/s |
|---|---|---|---|---|
| CoreWeave | — | — | $1.34 | 239 |
| Together | — | — | $2.02 | 166 |
| Decart | — | — | $1.35 | 129 |
| Fireworks | — | — | $1.71 | 111 |
| ModelRun | — | — | $1.38 | 110 |
| DeepInfra | — | — | $1.44 | 65 |
| Crusoe | — | — | $1.40 | 63 |
| Venice | — | — | $1.44 | 63 |
| Cloudflare | — | — | $1.71 | 56 |
| DigitalOcean | — | — | $1.37 | 52 |
| Chutes | — | — | $1.37 | 47 |
| SiliconFlow | — | — | $1.43 | 47 |
| Phala | — | — | $1.97 | 46 |
| Parasail | — | — | $1.44 | 45 |
| Inceptron | — | — | $1.30 | 37 |
| Moonshot AI · first party | — | — | $1.71 | 37 |
| Novita | — | — | $1.45 | 36 |
| StreamLake | — | — | $1.54 | 36 |
| Baidu · first party | — | — | $1.16 | 33 |
| AtlasCloud | — | — | $1.71 | 32 |
| BaseTen | — | — | $1.71 | — |
Per-host pricing and telemetry from OpenRouter. A dash means that host does not serve that model.