Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear. Sort by: Total Cost ↓.
408/408
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #136 | GLM 5.3 FlashX high | Z.ai | 0.8 | 0.7 | $0.210 | 2/3 | 8.06s |
| #162 | MiMo-V2.6-Flash low | Xiaomi | 0.8 | 0.7 | $0.200 | 2/3 | 11.0s |
| #160 | Qwen3.8 Flash Next xhigh | Qwen | 0.7 | 0.7 | ~$0.195 | 2/3 | 192.2s |
| #337 | Nemotron 3.5 Lightning medium | NVIDIA | 0.4 | 0.5 | $0.195 | 0/3 | 285.5s |
| #363 | GLM 4.7 Flash medium | Z.ai | 0.3 | 0.4 | $0.193 | 0/3 | 55.3s |
| #256 | MiMo-V2.6-Pro none | Xiaomi | 0.6 | 0.6 | $0.193 | 1/3 | 3.69s |
| #181 | Qwen3.5-Flash medium | Qwen | 0.4 | 0.7 | $0.184 | 0/3 | 58.9s |
| #227 | Gemini 3.5 Flash Lite low | 0.4 | 0.6 | $0.183 | 0/3 | 917ms | |
| #25 | Gemini 3.7 Flash low | 0.8 | 0.9 | $0.183 | 2/3 | 6.20s | |
| #275 | Seed 2.1 Turbo none | Bytedance Seed | 0.5 | 0.6 | $0.182 | 0/3 | 2.35s |
| #115 | Qwen3.7 Plus none | Qwen | 0.5 | 0.8 | $0.178 | 1/3 | 2.15s |
| #231 | Qwen3.5 Plus 2026-02-15 none | Qwen | 0.4 | 0.6 | $0.176 | 0/3 | 2.05s |
| #347 | Nemotron 3.5 Lightning low | NVIDIA | 0.4 | 0.5 | $0.175 | 0/3 | 224.4s |
| #273 | GPT-5.4 Mini none | OpenAI | 0.5 | 0.6 | $0.170 | 1/3 | 913ms |
| #308 | Nemotron 3.5 Lightning high | NVIDIA | 0.4 | 0.5 | $0.168 | 0/3 | 168.4s |