Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear. Sort by: Total Cost ↑.
404/404
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #301 | Laguna XS 2.1 none | Poolside | 0.4 | 0.5 | $0.020 | 0/3 | 623ms |
| #311 | Mercury 2.5 low | Inception | 0.5 | 0.5 | $0.020 | 1/3 | 972ms |
| #398 | Step 3.5 Flash none | Stepfun | 1.0 | 0.2 | $0.020 | 0/1 | 0ms |
| #267 | Ling 3.0 Flash high | Inclusionai | 0.7 | 0.6 | $0.021 | 1/3 | 14.9s |
| #351 | Mimo V2 Omni default | Xiaomi | 0.4 | 0.4 | $0.021 | 0/3 | 2.75s |
| #278 | Solar Pro 4 none | Upstage | 0.5 | 0.6 | $0.022 | 1/3 | 2.61s |
| #333 | Mercury 2.5 Preview default | Inception | 0.5 | 0.5 | $0.023 | 0/3 | 459ms |
| #319 | GPT-4o-mini default | OpenAI | 0.3 | 0.5 | $0.024 | 0/3 | 1.63s |
| #385 | MiMo-V2-Flash default | Xiaomi | 0.4 | 0.3 | $0.025 | 0/3 | 2.64s |
| #285 | GPT-6 Luna none | OpenAI | 0.5 | 0.5 | $0.026 | 1/3 | 8.04s |
| #350 | Ring 2.6 1t default | Inclusionai | 0.5 | 0.4 | $0.026 | 1/3 | 143.8s |
| #363 | Mercury 2.5 none | Inception | 0.4 | 0.4 | $0.026 | 0/3 | 12.5s |
| #332 | GLM 4.7 Flash none | Z.ai | 0.4 | 0.5 | $0.027 | 0/3 | 2.54s |
| #263 | Qwen3.7 Flash none | Qwen | 0.5 | 0.6 | $0.028 | 1/3 | 1.55s |
| #249 | Granite 4.2 8B medium | IBM Granite | 0.5 | 0.6 | $0.029 | 1/3 | 12.1s |