Quality
Known Good Index and the capability indices, side by side
Each index is min-max normalised across tracked models and scaled 0–100 from published evaluations. We do not run these evaluations.
| Metric | OpenAI: GPT-5 Mini | Z.ai: GLM 4.7 |
|---|---|---|
| Known Good Index | 60 | 68 |
| Coding Index | 50 | — |
| Mathematics Index | 93 | 83 |
| Science Index | 76 | 86 |
| Reasoning Index | 57 | — |
| Agentic Index | 50 | 35 |
| Openness Index | 5 | 60 |
Quality Breakdown
Every evaluation both models report a published score for
Bar = OpenAI: GPT-5 Mini · marker = Z.ai: GLM 4.7 · scores ingested from Epoch AI (CC-BY 4.0)
Right-hand figures read OpenAI: GPT-5 Mini / Z.ai: GLM 4.7. OpenAI: GPT-5 Mini leads on 2 of 3, Z.ai: GLM 4.7 on 1.
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena. Higher is better.
| Metric | OpenAI: GPT-5 Mini | Z.ai: GLM 4.7 |
|---|---|---|
| Text Arena | — | 1,436 |
| Text Arena (style controlled) | — | 1,442 |
Price & Cost
Published API pricing by token type
USD per 1M tokens. Blended is a 3:1 input:output mix. Live from OpenRouter, cross-checked against models.dev. Lower is better.
| Metric | OpenAI: GPT-5 Mini | Z.ai: GLM 4.7 |
|---|---|---|
| Input $/1M | $0.25 | $0.40 |
| Output $/1M | $2.00 | $1.75 |
| Blended $/1M | $0.69 | $0.74 |
| Cache read $/1M | $0.025 | $0.080 |
Context Window
Maximum input accepted and maximum output emitted
Provider-declared limits, from metadata via models.dev and OpenRouter. Higher is better.
| Metric | OpenAI: GPT-5 Mini | Z.ai: GLM 4.7 |
|---|---|---|
| Context window | 400k | 205k |
| Max output tokens | 128k | 131k |
Speed & Latency
Median throughput and time to first token across hosting providers
Medians across all tracked providers, derived from OpenRouter endpoint telemetry.
| Metric | OpenAI: GPT-5 Mini | Z.ai: GLM 4.7 |
|---|---|---|
| Output tokens/s | 100 | 66 |
| Time to first token | 3.3s | 2.1s |
Providers
Which hosts serve each model, and at what price and speed
| Provider | OpenAI: GPT-5 Mini blended | OpenAI: GPT-5 Mini tok/s | Z.ai: GLM 4.7 blended | Z.ai: GLM 4.7 tok/s |
|---|---|---|---|---|
| OpenAI · first party | $0.34 | 142 | — | — |
| Azure · first party | $0.69 | 78 | — | — |
| Cerebras | — | — | $2.38 | 401 |
| Google · first party | — | — | $1.00 | 140 |
| AtlasCloud | — | — | $0.85 | 54 |
| DeepInfra | — | — | $0.74 | 49 |
| Venice | — | — | $1.07 | 46 |
| StreamLake | — | — | $0.80 | 41 |
| Phala | — | — | $1.46 | 31 |
| Z.AI · first party | — | — | $1.00 | 30 |
| Novita | — | — | $0.90 | 27 |
Per-host pricing and telemetry from OpenRouter. A dash means that host does not serve that model.