AI BENCHY Failures
API error Failures
See which AI models run into API error most often, so you can spot reliability risks before choosing one. Sort by: Response Time (avg) ↑.
| Rank | Model | Company | API error Count | Score | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|
| #14 | Gemma 4 31B medium | 2 | 8.3 | 13/18 | 24.9s | |
| #43 | Qwen3.5-35B-A3B medium | Qwen | 1 | 7.4 | 10/18 | 44.5s |
| #32 | Qwen3.5-Flash medium | Qwen | 1 | 7.8 | 11/18 | 66.7s |