Open models, explained.
Which open model is best for your use case, what it costs, and how it compares to closed models.
Pick by use case
One index hides the fit. Choose what you are building; the answer and the evidence follow.
Which model is smartest overall?
$2.01 per task · 84% of the leader's score for 81% less
$0.25 per task · 93% of the open pick, 11% of the cost
$7.63 per task · 8.5 points ahead of the open pick
Score breakdown
| Model | Score | Cost / task | AA-Briefcase15% | GDPval-AA v210% | Terminal-Bench 4.010% | SciCode10% | Humanity's Last Exam10% | CritPt10% | GDP.pdf10% | Omniscience accuracy10% | Not guessing5% | AutomationBench-AA5% | AA-LCR v1.15% | GPQA Diamond |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
Claude Fable 5.1 | 53.4 | $7.63 | 58.1 | 63.2 | 52.0 | 63.1 | 59.1 | 29.7 | 26.2 | 67.2 | 27.4 | 59.4 | 85.3 | 93.7 |
GPT-6 Astra | 52.8 | $3.26 | 53.1 | 54.0 | 59.1 | 56.5 | 54.7 | 31.7 | 31.0 | 62.6 | 48.7 | 68.5 | 80.7 | 96.1 |
Claude Opus 5 | 50.7 | $5.86 | 57.2 | 61.8 | 49.0 | 56.4 | 54.9 | 29.1 | 21.6 | 60.9 | 39.2 | 56.6 | 79.3 | 93.2 |
Claude Fable 5 | 49.7 | $8.75 | 51.5 | 56.6 | 42.4 | 61.0 | 55.5 | 28.6 | 24.0 | 65.3 | 36.4 | 54.1 | 82.3 | 92.6 |
Muse Spark 1.3 | 48.2 | $1.60 | 54.5 | 60.2 | 33.3 | 58.8 | 48.7 | 24.9 | 26.6 | 43.6 | 67.1 | 57.9 | 83.0 | 93.5 |
GPT-5.6 Sol | 47.1 | $1.99 | 48.8 | 56.2 | 39.9 | 57.1 | 49.5 | 32.3 | 27.2 | 59.4 | 7.8 | 60.1 | 84.0 | 94.1 |
Qwen3.8 Max | 45.4 | $5.41 | 56.1 | 59.4 | 38.9 | 52.1 | 43.1 | 17.7 | 22.8 | 31.7 | 71.2 | 56.2 | 80.3 | 92.8 |
ZGLM-5.3 | 44.9 | $2.01 | 50.6 | 57.8 | 41.9 | 59.0 | 42.3 | 19.1 | 11.2 | 33.9 | 70.4 | 62.2 | 79.7 | 91.7 |
Grok 4.6 | 44.4 | $1.86 | 51.7 | 57.1 | 21.2 | 56.5 | 42.9 | 17.1 | 17.0 | 48.2 | 65.7 | 66.7 | 80.3 | 94.9 |
Kimi K3 | 43.8 | $2.00 | 49.4 | 52.6 | 12.6 | 59.5 | 46.9 | 23.4 | 22.0 | 47.6 | 46.8 | 58.3 | 88.7 | 93.5 |
Every model, ranked
The whole field on one index. Sort, filter, click to inspect.
Artificial Analysis Intelligence Index
Leaderboard
| Model | |||||
|---|---|---|---|---|---|
| 1 | Claude Fable 5.1 | 53 | 0.4 | $13.1K | |
| 2 | GPT-6 Astra | 53 | 1.0 | $5.3K | |
| 3 | Claude Opus 5 | 51 | 0.7 | $7.3K | |
| 4 | Claude Fable 5 | 50 | 0.4 | $11.2K | |
| 5 | Muse Spark 1.3 | 48 | 2.4 | $2.0K | |
| 6 | GPT-5.6 Sol | 47 | 1.4 | $3.5K | |
| 7 | Qwen3.8 Max | 45 | 0.9 | $4.9K | |
| 8 | Z GLM-5.3 | 45 | 1.8 | $2.5K | |
| 9 | Grok 4.6 | 44 | 1.9 | $2.4K | |
| 10 | Kimi K3 | 44 | 1.2 | $3.7K |
GLM-5.3
- $2.01
- $2.5K
- $1.40 / $4.40
- $0.26
- 228 t/s
- 1M
zai-org/GLM-5.3Run it The state of open, at a glance
Who is smartest, and who gives the most per dollar.
Intelligence
Value
GLM-5.3 is 8.5 points behind Claude Fable 5.1 on the Intelligence Index, holding 84% of the top score. In 2025 that gap was 9.3 points, and it has narrowed since.
Run the open frontier yourself
Same OpenAI-compatible API, open source weights you keep, at a fraction of closed-API cost. The snippet is pointed at GLM-5.3, the highest-scoring open model Together serves today and the open leader overall.