Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear. Sort by: Total Cost ↑.
404/404
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #260 | Qwen3.6 Flash none | Qwen | 0.5 | 0.6 | $0.139 | 1/3 | 1.79s |
| #35 | Claude Haiku 5.5 xhigh | Anthropic | 1.0 | 0.9 | $0.140 | 3/3 | 13.1s |
| #279 | GLM 5 none | Z.ai | 0.4 | 0.6 | $0.140 | 0/3 | 5.12s |
| #129 | Seed-2.0-Mini medium | Bytedance Seed | 0.6 | 0.8 | $0.140 | 1/3 | 220.5s |
| #175 | Gemma 4 31B medium | 0.4 | 0.7 | $0.144 | 0/3 | 219.8s | |
| #324 | Laguna S 2.1 high | Poolside | 0.5 | 0.5 | $0.145 | 1/3 | 271.4s |
| #225 | Qwen3.6 27B none | Qwen | 0.5 | 0.6 | $0.148 | 1/3 | 4.16s |
| #92 | DeepSeek V4 Pro 0423 high | DeepSeek | 0.6 | 0.8 | $0.149 | 1/3 | 243.0s |
| #236 | DeepSeek V4 Flash 0423 none | DeepSeek | 0.4 | 0.6 | $0.149 | 0/3 | 17.1s |
| #234 | DeepSeek V3.2 medium | DeepSeek | 0.6 | 0.6 | $0.152 | 1/3 | 248.7s |
| #173 | Gemini 3.1 Flash Lite medium | 0.5 | 0.7 | $0.153 | 1/3 | 3.81s | |
| #214 | Gemini 3 Flash Preview none | 0.5 | 0.6 | $0.155 | 1/3 | 1.80s | |
| #253 | GLM 5.3 FlashX low | Z.ai | 0.5 | 0.6 | $0.157 | 1/3 | 5.45s |
| #166 | Gemini 3.1 Flash Lite Preview medium | 0.5 | 0.7 | $0.158 | 1/3 | 4.09s | |
| #270 | GPT-5 Nano medium | OpenAI | 0.7 | 0.6 | $0.159 | 1/3 | 41.6s |