Every LLM benchmark, one honest table.
Scores, prices, and speed for 129 models — aggregated from public sources, with the origin and provenance of every number visible. Refreshed automatically, last 2h ago.
Frontier right now
Full leaderboard →| # | Model | Lab | Context | $/1M in · out | Intelligence | Coding | Agentic | Arena Elo |
|---|---|---|---|---|---|---|---|---|
| 1 | Claude Fable 5 | Anthropic | 1M | $10 · $50 | 59.9 | 76.5 | 52.8 | 1504 |
| 2 | GPT-5.6 Sol | OpenAI | 1.1M | $5 · $30 | 58.9 | 77.4 | 54.0 | 1486 |
| 3 | Kimi K3open | Moonshot AI | 1.0M | $3 · $15 | 57.1 | 76.2 | 50.1 | 1482 |
| 4 | Claude Opus 4.8 | Anthropic | 1M | $5 · $25 | 55.7 | 74.3 | 47.2 | 1472 |
| 5 | GPT-5.6 Terra | OpenAI | 1.1M | $2.5 · $15 | 55.0 | 76.7 | 47.4 | — |
| 6 | GPT-5.5 | OpenAI | 1.1M | $5 · $30 | 54.8 | 74.9 | 44.9 | 1490 |
| 7 | Grok 4.5 | xAI | 500K | $2 · $6 | 53.8 | 72.4 | 45.7 | 1478 |
| 8 | Claude Opus 4.7 | Anthropic | 1M | $5 · $25 | 53.5 | 73.6 | 44.4 | 1499 |
| 9 | Claude Sonnet 5 | Anthropic | 1M | $2 · $10 | 53.4 | 71.5 | 46.7 | 1460 |
| 10 | GPT-5.4 | OpenAI | 1.1M | $2.5 · $15 | 51.4 | 71.1 | 41.1 | 1499 |
| 11 | GPT-5.6 Luna | OpenAI | 1.1M | $1 · $6 | 51.2 | 71.4 | 45.6 | — |
| 12 | GLM-5.2 | Zhipu AI | 1M | $1.4 · $4.4 | 51.1 | 68.8 | 43.1 | — |
| 13 | Muse Spark 1.1 | Meta | 1M | $1.25 · $4.25 | 50.6 | 71.3 | 37.5 | 1491 |
| 14 | Gemini 3.5 Flash | 1.0M | $1.5 · $9 | 50.2 | 70.1 | 37.4 | 1490 | |
| 15 | Gemini 3.6 Flash | 1.0M | $1.5 · $7.5 | 50.1 | 69.2 | 38.7 | 1491 |
Intelligence, coding & agentic indices by Artificial Analysis; Arena Elo by LMArena (CC BY 4.0). Best value across model variants shown.
Intelligence vs. price
The frontier you actually pay for — up and to the left is better. Hover a point for details; click it to open the model. Free-tier models excluded.
API-onlyOpen weights
New models
- Gemini 3.6 FlashGoogleJul 21, 2026
- Gemini 3.5 Flash LiteGoogleJul 21, 2026
- Kimi K3Moonshot AIJul 16, 2026
- GPT-5.6 TerraOpenAIJul 9, 2026
- GPT-5.6OpenAIJul 9, 2026
- GPT-5.6 LunaOpenAIJul 9, 2026
- GPT-5.6 SolOpenAIJul 9, 2026
- Grok 4.5xAIJul 8, 2026
How to read the numbers
- independent — measured by an independent evaluator
- crowd — human preference votes (Elo)
- mirror — mirrored via an aggregator API
- vendor — self-reported by the model's vendor
Every score keeps a link to where it came from and when we saw it. Read the methodology.