Find your best-fit model
There is no best model, only a best fit. Move the three weights to match what your workload actually cares about and the table re-ranks as you drag. Everything runs in your browser against the same public data you can download — no account, no tracking, nothing sent anywhere.
Weights are relative — what matters is their proportion, not the totals.
Intelligence is the Known Good Index. Speed is median output tokens per second across tracked providers. Low cost is blended price, inverted so cheaper scores higher. Each is min-max normalised across every model that has a Known Good Index score, then combined with your weights.
It will not tell you a model is correct for your task. Aggregate scores hide the thing you actually care about, which is usually your own prompts on your own data — treat the shortlist as a shortlist and test the top two or three yourself.
It will not price your workload. We publish blended price per 1M tokens, not cost per task: cost per task needs token counts from an evaluation suite we do not run, so publishing it would mean inventing the inputs. See methodology.
It will not rank a model we cannot score. A model needs at least one science, one mathematics and one coding result before it appears here — 61 of 908 tracked models currently qualify.
| # | Model | Known Good Index | Output t/s | Blended $/M | Match |
|---|---|---|---|---|---|
| 1 |
|
98 | — | — | — |
| 2 |
|
96 | 104 | $4.50 | — |
| 3 |
|
95 | 119 | $3.38 | — |
| 4 |
|
94 | 52 | $10.00 | — |
| 5 |
|
93 | 40 | $2.21 | — |
| 6 |
|
93 | 45 | $0.54 | — |
| 7 |
|
93 | 66 | $1.16 | — |
| 8 |
|
91 | 55 | $1.07 | — |
| 9 |
|
90 | 46 | $2.34 | — |
| 10 |
|
90 | 82 | $5.62 | — |
| 11 |
|
88 | 37 | $10.00 | — |
| 12 |
|
88 | 53 | $1.48 | — |
| 13 |
|
83 | 89 | $1.12 | — |
| 14 |
|
83 | — | — | — |
| 15 |
|
82 | 132 | $1.93 | — |
| 16 |
|
80 | 54 | $4.81 | — |
| 17 |
|
80 | 41 | $6.00 | — |
| 18 |
|
78 | 47 | $0.73 | — |
| 19 |
|
77 | 37 | $1.35 | — |
| 20 |
|
74 | 62 | $3.44 | — |