Coding Ranking
See which AI models perform best on Coding, which ones stay reliable, and where the biggest gaps appear. Sort by: Tests Correct ↑.
408/408
Filter models
No models match the current search and filters.
| Rank | Model | Company | Coding Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #290 | GLM 5V Turbo medium | Z.ai | 0.6 | 0.5 | $0.457 | 1/3 | 63.4s |
| #293 | Qwen3.8 27B none | Qwen | 0.5 | 0.5 | ~$0.010 | 1/3 | 1.19s |
| #298 | Ember-1 none | Fireworks | 0.5 | 0.5 | $0.786 | 1/3 | 1.94s |
| #299 | Hy3 preview medium | Tencent | 0.5 | 0.5 | $0.049 | 1/3 | 31.4s |
| #301 | Granite 4.2 8B low | IBM Granite | 0.5 | 0.5 | $0.036 | 1/3 | 13.0s |
| #302 | Kimi K2.5 none | Moonshot AI | 0.5 | 0.5 | $0.262 | 1/3 | 24.6s |
| #305 | Qwen3.6 35B A3B none | Qwen | 0.5 | 0.5 | $0.105 | 1/3 | 8.77s |
| #307 | Ling 3.0 Flash none | Inclusionai | 0.5 | 0.5 | $0.004 | 1/3 | 6.39s |
| #310 | Mimo V2 PRO medium | Xiaomi | 0.6 | 0.5 | $0.333 | 1/3 | 94.2s |
| #311 | MiMo-V2-Flash medium | Xiaomi | 0.6 | 0.5 | $0.043 | 1/3 | 10.7s |
| #314 | Mercury 2.5 low | Inception | 0.5 | 0.5 | $0.020 | 1/3 | 972ms |
| #318 | Laguna S 2.1 medium | Poolside | 0.5 | 0.5 | $0.092 | 1/3 | 159.0s |
| #319 | Mercury 2.5 Preview low | Inception | 0.5 | 0.5 | $0.015 | 1/3 | 972ms |
| #320 | MiniMax M2.7 medium | Minimax | 0.6 | 0.5 | $0.326 | 1/3 | 101.9s |
| #327 | Laguna S 2.1 high | Poolside | 0.5 | 0.5 | $0.145 | 1/3 | 271.4s |