Quality
Known Good Index and the capability indices, side by side
Each index is min-max normalised across tracked models and scaled 0–100 from published evaluations. We do not run these evaluations.
| Metric | OpenAI: o3 | Z.ai: GLM 5.1 |
|---|---|---|
| Known Good Index | 67 | 88 |
| Mathematics Index | 92 | 92 |
| Science Index | 84 | 89 |
| Reasoning Index | 62 | — |
| Agentic Index | 60 | 82 |
| Openness Index | 5 | 5 |
Quality Breakdown
Every evaluation both models report a published score for
Bar = OpenAI: o3 · marker = Z.ai: GLM 5.1 · scores ingested from Epoch AI (CC-BY 4.0)
Right-hand figures read OpenAI: o3 / Z.ai: GLM 5.1. OpenAI: o3 leads on 0 of 3, Z.ai: GLM 5.1 on 3.
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena. Higher is better.
| Metric | OpenAI: o3 | Z.ai: GLM 5.1 |
|---|---|---|
| Text Arena | — | 1,465 |
| Text Arena (style controlled) | — | 1,470 |
Price & Cost
Published API pricing by token type
USD per 1M tokens. Blended is a 3:1 input:output mix. Live from OpenRouter, cross-checked against models.dev. Lower is better.
| Metric | OpenAI: o3 | Z.ai: GLM 5.1 |
|---|---|---|
| Input $/1M | $2.00 | $0.97 |
| Output $/1M | $8.00 | $3.04 |
| Blended $/1M | $3.50 | $1.48 |
| Cache read $/1M | $0.500 | $0.179 |
Context Window
Maximum input accepted and maximum output emitted
Provider-declared limits, from metadata via models.dev and OpenRouter. Higher is better.
| Metric | OpenAI: o3 | Z.ai: GLM 5.1 |
|---|---|---|
| Context window | 200k | 205k |
| Max output tokens | 100k | 128k |
Speed & Latency
Median throughput and time to first token across hosting providers
Medians across all tracked providers, derived from OpenRouter endpoint telemetry.
| Metric | OpenAI: o3 | Z.ai: GLM 5.1 |
|---|---|---|
| Output tokens/s | 88 | 53 |
| Time to first token | 2.7s | 1.9s |
Providers
Which hosts serve each model, and at what price and speed
| Provider | OpenAI: o3 blended | OpenAI: o3 tok/s | Z.ai: GLM 5.1 blended | Z.ai: GLM 5.1 tok/s |
|---|---|---|---|---|
| OpenAI · first party | $3.50 | 72 | — | — |
| CoreWeave | — | — | $2.15 | 120 |
| Friendli | — | — | $2.15 | 109 |
| Crusoe | — | — | $2.00 | 91 |
| Fireworks | — | — | $2.15 | 90 |
| Venice | — | — | $2.37 | 80 |
| GMICloud | — | — | $1.50 | 77 |
| Parasail | — | — | $2.15 | 71 |
| DeepInfra | — | — | $1.66 | 66 |
| SiliconFlow | — | — | $1.83 | 62 |
| Novita | — | — | $2.13 | 52 |
| Baidu · first party | — | — | $1.50 | 51 |
| Wafer | — | — | $1.55 | 50 |
| Chutes | — | — | $1.50 | 48 |
| AtlasCloud | — | — | $1.94 | 42 |
| Z.AI · first party | — | — | $2.15 | 37 |
| Nebius | — | — | $2.15 | 26 |
| StreamLake | — | — | $1.48 | 25 |
| DigitalOcean | — | — | $1.81 | 21 |
| Phala | — | — | $1.96 | — |
Per-host pricing and telemetry from OpenRouter. A dash means that host does not serve that model.