Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear. Sort by: Total Cost ↑.
404/404
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #241 | DeepSeek V4 Flash 0731 none | DeepSeek | 0.4 | 0.6 | $0.053 | 0/3 | 1.72s |
| #322 | Inkling Small none | Thinkingmachines | 0.4 | 0.5 | $0.054 | 0/3 | 759ms |
| #303 | Inkling Small low | Thinkingmachines | 0.5 | 0.5 | $0.055 | 0/3 | 3.70s |
| #142 | GLM 5.3 Flash high | Z.ai | 0.8 | 0.7 | $0.056 | 2/3 | 19.6s |
| #381 | Grok 4.20 none | X AI | 0.1 | 0.3 | $0.057 | 0/1 | 1.22s |
| #293 | MiMo-V2.6-Flash none | Xiaomi | 0.4 | 0.5 | $0.060 | 0/3 | 744ms |
| #64 | Claude Haiku 5.5 low | Anthropic | 0.8 | 0.9 | $0.061 | 2/3 | 7.30s |
| #154 | Qwen3.7 Flash high | Qwen | 0.8 | 0.7 | $0.063 | 2/3 | 58.8s |
| #120 | Claude Haiku 5.5 medium | Anthropic | 0.6 | 0.8 | $0.063 | 1/3 | 7.16s |
| #382 | Command A+ medium | Cohere | 0.3 | 0.3 | $0.064 | 0/3 | 93.1s |
| #132 | GPT-5.6 Luna medium | OpenAI | 0.5 | 0.7 | $0.065 | 1/3 | 10.4s |
| #323 | Qwen3 Coder Next default | Qwen | 0.5 | 0.5 | $0.067 | 0/3 | 2.22s |
| #185 | Solar Pro 4 xhigh | Upstage | 0.6 | 0.7 | $0.067 | 1/3 | 292.2s |
| #340 | GPT-5.4 Nano none | OpenAI | 0.5 | 0.5 | $0.068 | 0/3 | 2.22s |
| #368 | Grok 4.1 Fast medium | X AI | 0.8 | 0.4 | $0.069 | 0/1 | 23.6s |