Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear. Sort by: Metric ↑.
404/404
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #248 | Qwen3.5 Plus 2026-04-20 none | Qwen | 0.4 | 0.6 | $0.216 | 0/3 | 1.69s |
| #328 | Qwen3.5-9B none | Qwen | 0.4 | 0.5 | $0.041 | 0/3 | 5.60s |
| #362 | GLM 5 Turbo none | Z.ai | 0.4 | 0.4 | $0.047 | 0/3 | 2.41s |
| #202 | Step 3.7 Flash high | Stepfun | 0.4 | 0.7 | $1.478 | 0/3 | 206.2s |
| #279 | GLM 5 none | Z.ai | 0.4 | 0.6 | $0.140 | 0/3 | 5.12s |
| #344 | Nemotron 3.5 Lightning low | NVIDIA | 0.4 | 0.5 | $0.175 | 0/3 | 224.4s |
| #378 | Elephant Alpha default | Openrouter | 0.4 | 0.4 | $0.000 | 0/3 | 1.39s |
| #370 | Ling 3.0 Tiny default | Inclusionai | 0.4 | 0.4 | $0.000 | 0/3 | 1.19s |
| #274 | Dots 3 Note Preview medium | Dots Studio | 0.4 | 0.6 | $0.000 | 0/3 | 309.1s |
| #236 | DeepSeek V4 Flash 0423 none | DeepSeek | 0.4 | 0.6 | $0.149 | 0/3 | 17.1s |
| #226 | Gemini 3.5 Flash Lite low | 0.4 | 0.6 | $0.183 | 0/3 | 917ms | |
| #230 | Qwen3.5 Plus 2026-02-15 none | Qwen | 0.4 | 0.6 | $0.176 | 0/3 | 2.05s |
| #385 | MiMo-V2-Flash default | Xiaomi | 0.4 | 0.3 | $0.025 | 0/3 | 2.64s |
| #175 | Gemma 4 31B medium | 0.4 | 0.7 | $0.144 | 0/3 | 219.8s | |
| #251 | Apodex 1.1 Mini medium | Apodex | 0.4 | 0.6 | $0.000 | 0/3 | 51.0s |