Quality
Known Good Index and the capability indices, side by side
Each index is min-max normalised across tracked models and scaled 0–100 from published evaluations. We do not run these evaluations.
| Metric | OpenAI: o1 | Qwen: Qwen3.6 Max Preview |
|---|---|---|
| Known Good Index | 63 | 90 |
| Mathematics Index | 85 | 91 |
| Science Index | 78 | 93 |
| Reasoning Index | 62 | — |
| Agentic Index | — | 87 |
| Openness Index | 5 | 5 |
Quality Breakdown
Every evaluation both models report a published score for
Bar = OpenAI: o1 · marker = Qwen: Qwen3.6 Max Preview · scores ingested from Epoch AI (CC-BY 4.0)
Right-hand figures read OpenAI: o1 / Qwen: Qwen3.6 Max Preview. OpenAI: o1 leads on 0 of 2, Qwen: Qwen3.6 Max Preview on 2.
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena. Higher is better.
| Metric | OpenAI: o1 | Qwen: Qwen3.6 Max Preview |
|---|---|---|
| Text Arena | — | 1,456 |
| Text Arena (style controlled) | — | 1,460 |
Price & Cost
Published API pricing by token type
USD per 1M tokens. Blended is a 3:1 input:output mix. Live from OpenRouter, cross-checked against models.dev. Lower is better.
| Metric | OpenAI: o1 | Qwen: Qwen3.6 Max Preview |
|---|---|---|
| Input $/1M | $15.00 | $1.04 |
| Output $/1M | $60.00 | $6.24 |
| Blended $/1M | $26.25 | $2.34 |
| Cache read $/1M | $7.500 | $0.130 |
Context Window
Maximum input accepted and maximum output emitted
Provider-declared limits, from metadata via models.dev and OpenRouter. Higher is better.
| Metric | OpenAI: o1 | Qwen: Qwen3.6 Max Preview |
|---|---|---|
| Context window | 200k | 262k |
| Max output tokens | 100k | 66k |
Speed & Latency
Median throughput and time to first token across hosting providers
Medians across all tracked providers, derived from OpenRouter endpoint telemetry.
| Metric | OpenAI: o1 | Qwen: Qwen3.6 Max Preview |
|---|---|---|
| Output tokens/s | 24 | 46 |
| Time to first token | 1.4s | 1.4s |
Providers
Which hosts serve each model, and at what price and speed
| Provider | OpenAI: o1 blended | OpenAI: o1 tok/s | Qwen: Qwen3.6 Max Preview blended | Qwen: Qwen3.6 Max Preview tok/s |
|---|---|---|---|---|
| OpenAI · first party | $26.25 | 13 | — | — |
| Alibaba · first party | — | — | $2.34 | 46 |
Per-host pricing and telemetry from OpenRouter. A dash means that host does not serve that model.