API error Failures
See which AI models run into API error most often, so you can spot reliability risks before choosing one. Sort by: Score ↑.
Categories
68/68
Filter models
No models match the current search and filters.
| Rank | Model | Company | API error Count | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #61 | Qwen3.5 Plus 2026-02-15 medium | Qwen | 1 | 7.5 | $0.437 | 14/22 | 89.2s |
| #56 | Kimi K2.7 Code medium | Moonshot AI | 1 | 7.5 | $0.740 | 12/22 | 84.2s |
| #55 | Nemotron 3 Ultra medium | NVIDIA | 1 | 7.5 | $0.774 | 13/22 | 32.2s |
| #50 | DeepSeek V4 Pro high | DeepSeek | 1 | 7.7 | $0.200 | 10/22 | 79.1s |
| #41 | Qwen3.6 Plus medium | Qwen | 1 | 7.8 | $0.405 | 15/22 | 43.1s |
| #37 | Kimi K3 max | Moonshot AI | 2 | 8.0 | $3.112 | 16/22 | 122.5s |
| #36 | Inkling medium | Thinkingmachines | 1 | 8.0 | $0.391 | 15/22 | 16.2s |
| #30 | Muse Spark 1.1 high | Meta | 1 | 8.1 | $1.694 | 12/22 | 31.5s |