Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear.
408/408
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #302 | Kimi K2.5 none | Moonshot AI | 0.5 | 0.5 | $0.262 | 1/3 | 24.6s |
| #305 | Qwen3.6 35B A3B none | Qwen | 0.5 | 0.5 | $0.105 | 1/3 | 8.77s |
| #307 | Ling 3.0 Flash none | Inclusionai | 0.5 | 0.5 | $0.004 | 1/3 | 6.39s |
| #314 | Mercury 2.5 low | Inception | 0.5 | 0.5 | $0.020 | 1/3 | 972ms |
| #319 | Mercury 2.5 Preview low | Inception | 0.5 | 0.5 | $0.015 | 1/3 | 972ms |
| #330 | MiMo-V2.5 none | Xiaomi | 0.5 | 0.5 | $0.049 | 1/3 | 3.24s |
| #345 | GLM 5V Turbo none | Z.ai | 0.5 | 0.5 | $0.052 | 1/3 | 3.13s |
| #349 | Mimo V2 PRO default | Xiaomi | 0.5 | 0.5 | $0.045 | 1/3 | 2.65s |
| #358 | Granite 4.2 8B none | IBM Granite | 0.5 | 0.4 | $0.037 | 1/3 | 105.6s |
| #96 | GPT-5.6 Luna high | OpenAI | 0.5 | 0.8 | $0.118 | 1/3 | 15.6s |
| #190 | Solar Pro 4 low | Upstage | 0.5 | 0.7 | $0.101 | 1/3 | 259.6s |
| #216 | Solar Pro 4 medium | Upstage | 0.5 | 0.6 | $0.098 | 1/3 | 310.4s |
| #255 | GLM 5.3 FlashX low | Z.ai | 0.5 | 0.6 | $0.157 | 1/3 | 5.45s |
| #262 | Qwen3.6 Flash none | Qwen | 0.5 | 0.6 | $0.139 | 1/3 | 1.79s |
| #348 | Owl Alpha medium | Openrouter | 0.5 | 0.5 | $0.000 | 1/3 | 18.7s |