Companies, each fielding its best model
Each company is represented by its strongest model, row intact.
- SOTA
- 13models
- Scores above the executioner
- No. 1 · AA
- Claude Opus 5.5
- 57.6 · $8.00
- Executioner · price
- Claude Haiku 5.5
- $0.20
- Cheapest SOTA
- MiMo V2.6 Pro
- 46.3 · $0.54
Ranking · AA 28–58 · ARENA 1416–1534Every model that outscores the executioner, with the executioner itself as the last row for reference. Each column is one board’s own score, side by side and never combined; a dash means that board does not list the model. Click a column head to sort by it. The rank is that board’s own, and hovering shows every board’s. A model marked Legacy is no longer listed as current by its source, and its row shows the best setting still listed.
- 1scores from Claude Opus 5.557.61507
- 2scores from GPT 6 Astra52.71475
- 3scores from Gemini 4 Argon52.61525
- 4scores from Muse Spark 1.348.11494
- 5scores from Grok 4.746.41443
- 6scores from MiMo V2.6 Pro46.31480
- 7scores from Qwen 3.8 Max45.41483
- 8scores from GLM 5.344.81478
- 9scores from Step 543.71447
- 1043.61488
- 11NewExecutioner43.4—
Kill line · score × price The executioner is the best value on the board: every model that costs more and scores no higher is killed.Each dot is a model: further right costs more, higher up scores more. Where the dashed lines cross stands the executioner. Above it is SOTA, the top tier; below and to the left is cheaper and weaker, the low-cost picks.65 of 102 models are killed. The faint line joins the models nothing beats on both price and score. Legacy models are left off the chart. Select a model to open its details below.How the kill line is drawn →Who held it before →
The executioner is the best value on the board: every model that costs more and scores no higher is killed. How the kill line is drawn →Who held it before →
Kill lineFrontier: no model beats these on price and score at onceSOTALow-costKilledHollow dot or faint cross — estimated by the source
Claude Opus 5.5
Scores on each board Each rank is the board’s own: the rank it publishes, or the model’s place in the board’s own order where it publishes none. The bar shows how much of that board the model ranks ahead of. Scores from different boards cannot be compared.
- Artificial Analysis LLM Leaderboard57.6No. 1 of 375
- Arena — Agent0.143No. 1 of 52
- Artificial Analysis Coding Agents0.660No. 2 of 24
- Epoch AI Benchmarking Hub91.2%No. 3 of 82
- Terminal-Bench 4.064.8%No. 1 of 35
- Arena — Code1813No. 1 of 142
- Arena — Text1507No. 2 of 414
- Arena — Vision1293No. 11 of 165
Scores at each effort setting The same model at each effort setting. A bar runs from zero to that board’s best score, so read each board on its own and never across boards.
Artificial Analysis LLM Leaderboardbest at max
- low42.3
- medium51.2
- high53.6
- xhigh56.0
- max57.6
Tested at one setting only
- Arena — Agent · agenthigh0.143
- Artificial Analysis Coding Agents · Claude Codemax0.660
- Epoch AI Benchmarking Hubmax91.2%
- Terminal-Bench 4.0 · Claude Codemax64.8%
- Arena — Codemax1813
- Arena — Texthigh1507
- Arena — Visionhigh1293
Every company
-
OpenAI 46 with scores · 2 SOTA
-
Alibaba 35 with scores · 1 SOTA
-
Google 35 with scores · 1 SOTA
-
Anthropic 26 with scores · 5 SOTA
-
Mistral 26 with scores -
DeepSeek 18 with scores
-
SpaceXAI 18 with scores · 1 SOTA
-
Z AI 17 with scores · 1 SOTA
-
Meta 15 with scores · 1 SOTA
-
NVIDIA 15 with scores
-
InclusionAI 13 with scores -
Tencent 11 with scores
-
Allen Institute for AI 10 with scores
-
IBM 10 with scores
-
Amazon 8 with scores
-
Cohere 8 with scores
-
Upstage 8 with scores
-
AI21 Labs 7 with scores
-
Kimi 7 with scores · 1 SOTA -
MiniMax 7 with scores
-
StepFun 7 with scores · 1 SOTA
-
Xiaomi 7 with scores · 1 SOTA
-
Liquid AI 6 with scores
-
Microsoft 6 with scores
-
Nous Research 5 with scores -
Perplexity 5 with scores -
Baidu 4 with scores
-
ByteDance 4 with scores
-
LG AI Research 4 with scores -
MBZUAI Institute of Foundation Models 4 with scores
-
China Mobile 3 with scores -
Inception 3 with scores
-
LongCat 3 with scores
-
OpenBMB 3 with scores
-
AI9Stars 2 with scores
-
Arcee AI 2 with scores
-
Databricks 2 with scores -
KwaiKAT 2 with scores
-
Motif Technologies 2 with scores
-
Multiverse Computing 2 with scores
- Poolside 2 with scores
-
Reka AI 2 with scores
-
Sarvam 2 with scores
-
ServiceNow 2 with scores
-
TII UAE 2 with scores
-
Thinking Machines 2 with scores
-
Celeris 1 with scores
-
Deep Cogito 1 with scores - Korea Telecom 1 with scores
-
Nanbeige 1 with scores -
Naver 1 with scores -
Nex AGI 1 with scores
-
OpenChat 1 with scores -
Prime Intellect 1 with scores
-
SK Telecom 1 with scores
-
Snowflake 1 with scores
-
Swiss AI Initiative 1 with scores -
Trillion Labs 1 with scores