Speed
How fast output arrives — sustained throughput and time to the first token, which trade off against each other and against cost. Measured on a provider's serving of the model, so it describes an endpoint rather than the weights. Ordered by how many of these benchmarks put a model in their top 10, then by its best placing. Each column keeps its own ranking; nothing is averaged. A dash means the board does not list that model — not a score of zero. Entries not yet matched to a model are left out here; they still appear on their own board.
| Model | In top 10 | Artificial Analysis LLM Leaderboard Output Speed | Artificial Analysis LLM Leaderboard Time To First Token |
|---|---|---|---|
| | 2/2 | #1 1523/s | #6 0.56s |
| | 2/2 | #10 307/s | #7 0.57s |
| | 1/2 | #12 289/s | #1 0.32s |
| | 1/2 | #2 818/s | #135 3.21s |
| | 1/2 | #28 208/s | #2 0.46s |
| | 1/2 | #3 456/s | #148 6.51s |
| | 1/2 | #95 90/s | #3 0.47s |
| | 1/2 | #4 365/s | #151 8.96s |
| | 1/2 | #35 175/s | #4 0.48s |
| | 1/2 | #5 339/s | #60 1.21s |
| | 1/2 | #25 215/s | #5 0.53s |
| | 1/2 | #6 334/s | #116 2.50s |
| | 1/2 | #7 325/s | #154 12.05s |
| | 1/2 | #8 323/s | #146 5.54s |
| | 1/2 | #23 222/s | #8 0.59s |
| | 1/2 | #9 321/s | #80 1.70s |
| | 1/2 | #84 103/s | #9 0.59s |
| | 1/2 | #145 49/s | #10 0.60s |