Quality
Known Good Index and the capability indices, side by side
Each index is min-max normalised across tracked models and scaled 0–100 from published evaluations. We do not run these evaluations.
| Metric | Z.ai: GLM 4.7 | Z.ai: GLM 5.2 |
|---|---|---|
| Known Good Index | 68 | 91 |
| Mathematics Index | 83 | 86 |
| Science Index | 86 | 97 |
| Agentic Index | 35 | 91 |
| Openness Index | 60 | 60 |
Quality Breakdown
Every evaluation both models report a published score for
Bar = Z.ai: GLM 4.7 · marker = Z.ai: GLM 5.2 · scores ingested from Epoch AI (CC-BY 4.0)
Right-hand figures read Z.ai: GLM 4.7 / Z.ai: GLM 5.2. Z.ai: GLM 4.7 leads on 0 of 2, Z.ai: GLM 5.2 on 2.
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena. Higher is better.
| Metric | Z.ai: GLM 4.7 | Z.ai: GLM 5.2 |
|---|---|---|
| Text Arena | 1,436 | 1,464 |
| Text Arena (style controlled) | 1,442 | 1,469 |
Price & Cost
Published API pricing by token type
USD per 1M tokens. Blended is a 3:1 input:output mix. Live from OpenRouter, cross-checked against models.dev. Lower is better.
| Metric | Z.ai: GLM 4.7 | Z.ai: GLM 5.2 |
|---|---|---|
| Input $/1M | $0.40 | $0.73 |
| Output $/1M | $1.75 | $2.28 |
| Blended $/1M | $0.74 | $1.12 |
| Cache read $/1M | $0.080 | $0.135 |
Context Window
Maximum input accepted and maximum output emitted
Provider-declared limits, from metadata via models.dev and OpenRouter. Higher is better.
| Metric | Z.ai: GLM 4.7 | Z.ai: GLM 5.2 |
|---|---|---|
| Context window | 205k | 1.05M |
| Max output tokens | 131k | 131k |
Speed & Latency
Median throughput and time to first token across hosting providers
Medians across all tracked providers, derived from OpenRouter endpoint telemetry.
| Metric | Z.ai: GLM 4.7 | Z.ai: GLM 5.2 |
|---|---|---|
| Output tokens/s | 66 | 55 |
| Time to first token | 2.1s | 1.7s |
Providers
Which hosts serve each model, and at what price and speed
| Provider | Z.ai: GLM 4.7 blended | Z.ai: GLM 4.7 tok/s | Z.ai: GLM 5.2 blended | Z.ai: GLM 5.2 tok/s |
|---|---|---|---|---|
| Google · first party | $1.00 | 147 | — | — |
| Cerebras | $2.38 | 132 | — | — |
| DeepInfra | $0.74 | 49 | $1.45 | 49 |
| Venice | $1.07 | 49 | $2.15 | 39 |
| AtlasCloud | $0.85 | 47 | $1.94 | 38 |
| StreamLake | $0.80 | 37 | $1.11 | 37 |
| Z.AI · first party | $1.00 | 33 | $2.15 | 49 |
| Phala | $1.46 | 30 | $2.15 | 46 |
| Novita | $0.90 | 29 | $1.12 | 37 |
| CoreWeave | — | — | $1.18 | 181 |
| BaseTen | — | — | $2.15 | 119 |
| Wafer | — | — | $2.15 | 103 |
| Ionstream | — | — | $2.15 | 101 |
| Friendli | — | — | $2.15 | 80 |
| Cloudflare | — | — | $2.15 | 74 |
| Alibaba · first party | — | — | $1.27 | 69 |
| Fireworks | — | — | $2.15 | 65 |
| Decart | — | — | $1.52 | 60 |
| Chutes | — | — | $1.93 | 56 |
| Parasail | — | — | $2.15 | 56 |
| SiliconFlow | — | — | $2.00 | 48 |
| Together | — | — | $2.15 | 47 |
| AkashML | — | — | $1.18 | 46 |
| Morph | — | — | $1.85 | 43 |
| Baidu · first party | — | — | $1.12 | 40 |
| GMICloud | — | — | $1.42 | 39 |
| Inceptron | — | — | $1.43 | 29 |
| Io Net | — | — | $4.40 | 29 |
| Sail Research | — | — | $1.72 | 20 |
| Ambient | — | — | $1.89 | 14 |
| DigitalOcean | — | — | $1.89 | 13 |
Per-host pricing and telemetry from OpenRouter. A dash means that host does not serve that model.