Coding: API error
Coding
API error
See which AI models are most likely to hit API error on Coding, so you can spot weak points faster. Sort by: Tests Correct ↓.
Failure Reasons
35/35
Filter models
No models match the current search and filters.
| Rank | Model | Company | API error Count | Category Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #54 | DeepSeek V4 Flash 0731 high | DeepSeek | 1 | 6.3 | $0.134 | 1/3 | 252.7s |
| #63 | Qwen3.6 Plus medium | Qwen | 1 | 6.1 | $0.418 | 1/3 | 153.1s |
| #73 | Seed-2.0-Code low | Bytedance Seed | 1 | 7.0 | $0.767 | 1/3 | 109.7s |
| #75 | DeepSeek V4 Pro high | DeepSeek | 1 | 6.3 | $0.761 | 1/3 | 243.0s |
| #86 | Qwen3.5 Plus 2026-02-15 medium | Qwen | 1 | 6.6 | $0.459 | 1/3 | 180.7s |
| #126 | Seed-2.0-Code high | Bytedance Seed | 1 | 6.2 | $1.273 | 1/3 | 266.7s |
| #135 | MiMo-V2.5-Pro medium | Xiaomi | 1 | 6.2 | $0.221 | 1/3 | 92.1s |
| #140 | LongCat 2.0 low | Meituan | 1 | 6.6 | $0.412 | 1/3 | 479.3s |
| #154 | Hy3 preview medium | Tencent | 2 | 5.3 | $0.049 | 1/3 | 31.4s |
| #162 | Ring-2.6-1T medium | Inclusionai | 2 | 5.3 | $0.102 | 1/3 | 59.6s |
| #163 | Mimo V2 PRO medium | Xiaomi | 1 | 6.0 | $0.333 | 1/3 | 94.2s |
| #181 | Trinity Large Thinking low | Arcee AI | 2 | 5.3 | $0.658 | 1/3 | 344.4s |
| #192 | Hy3 preview high | Tencent | 2 | 5.3 | $0.135 | 1/3 | 99.8s |
| #208 | Owl Alpha medium | Openrouter | 1 | 5.4 | $0.000 | 1/3 | 18.7s |
| #210 | Mimo V2 PRO none | Xiaomi | 1 | 5.5 | $0.045 | 1/3 | 2.65s |