Qwen 3.8 Max vs
Muse Spark 1.3
Both are SOTA right now. Each row is one board’s call, and the marked figure is the one that board places higher. No row is added to another, so there is no overall winner: the boards measure different things, and where they split, that is the finding.Each rank is the board’s own: the rank it publishes, or the model’s place in the board’s own order where it publishes none. Scores from different boards cannot be compared.
| Board | Qwen 3.8 Max | Muse Spark 1.3 |
|---|---|---|
| Artificial Analysis LLM Leaderboard Artificial Analysis Intelligence Index | 45.4 No. 14 of 375 | 48.1 No. 9 of 375 |
| Arena — Agent Overall | 0.023 No. 23 of 52 | 0.040 No. 15 of 52 |
| Artificial Analysis Coding Agents Artificial Analysis Coding Agent Index | 0.433 No. 18 of 24 | 0.543 No. 13 of 24 |
| Epoch AI Benchmarking Hub FrontierMath Tiers 1–3 | 74.7% No. 19 of 82 | 74.4% No. 20 of 82 |
| Terminal-Bench 4.0 Terminal-Bench 4.0 | 27.0% No. 25 of 35 | 14.6% No. 32 of 35 |
| Arena — Code Overall | 1674 No. 10 of 142 | 1657 No. 13 of 142 |
| Arena — Text Overall (style control) | 1483 No. 22 of 414 | 1494 No. 9 of 414 |
| Arena — Vision Overall (style control) | 1301 No. 2 of 165 | 1290 No. 12 of 165 |
At a glance
| Developer | Alibaba | Meta |
|---|---|---|
| Weights | open | closed |
| Released | 2026-08-03 | — |
| Price Blended Price, USD per 1M tokens (3:1 input:output) | $3.00 | $2.00 |
| Value: where it sits on the kill line | SOTA | SOTA |
Qwen 3.8 Max →Muse Spark 1.3 → How the kill line is drawn →