API error Failures
See which AI models run into API error most often, so you can spot reliability risks before choosing one. Sort by: Tests Correct ↑.
Categories
68/68
Filter models
No models match the current search and filters.
| Rank | Model | Company | API error Count | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #163 | Mimo V2 Omni none | Xiaomi | 1 | 5.5 | $0.021 | 8/21 | 2.44s |
| #143 | North Mini Code medium | Cohere | 1 | 5.9 | $0.000 | 9/22 | 137.1s |
| #185 | Ring-2.6-1T none | Inclusionai | 6 | 4.8 | $0.026 | 9/22 | 55.1s |
| #187 | Grok 4.20 Multi Agent Beta medium | X AI | 2 | 4.8 | $5.599 | 8/18 | 9.69s |
| #190 | Hunter Alpha medium | OpenRouter | 1 | 4.7 | $0.000 | 8/18 | 10.3s |
| #50 | DeepSeek V4 Pro high | DeepSeek | 1 | 7.7 | $0.200 | 10/22 | 79.1s |
| #96 | LongCat 2.0 low | Meituan | 1 | 6.7 | $0.391 | 10/22 | 100.3s |
| #121 | Gemma 4 31B none | 2 | 6.2 | $0.021 | 10/22 | 5.34s | |
| #181 | Qwen3.6 Plus Preview medium | Qwen | 8 | 4.9 | $0.000 | 9/19 | 15.2s |
| #192 | Laguna M.1 medium | Poolside | 4 | 4.7 | $0.033 | 9/19 | 14.7s |
| #140 | Mimo V2 Omni medium | Xiaomi | 1 | 5.9 | $0.683 | 10/21 | 41.2s |
| #159 | Hy3 preview low | Tencent | 7 | 5.5 | $0.015 | 10/21 | 24.6s |
| #66 | KAT-Coder-Pro V2.5 low | Kwaipilot | 1 | 7.4 | $0.387 | 11/22 | 19.5s |
| #80 | DeepSeek V3.2 medium | DeepSeek | 2 | 7.0 | $0.078 | 11/22 | 68.6s |
| #85 | KAT-Coder-Pro V2.5 medium | Kwaipilot | 1 | 6.9 | $0.467 | 11/22 | 24.0s |