AI BENCHY Failures
Timed out Failures
See which AI models run into Timed out most often, so you can spot reliability risks before choosing one. Sort by: Tests Correct ↑.
| Rank | Model | Company | Timed out Count | Score | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|
| #37 | Gemma 4 26B A4B medium | 2 | 7.6 | 14/21 | 63.4s | |
| #17 | GLM 5 medium | Z.ai | 1 | 8.3 | 15/21 | 33.5s |
| #18 | Qwen3.7 Plus medium | Qwen | 1 | 8.2 | 15/21 | 38.9s |
| #11 | Claude Opus 4.7 medium | Anthropic | 1 | 8.7 | 17/21 | 4.73s |