No answer Failures
See which AI models run into No answer most often, so you can spot reliability risks before choosing one.
119/119
Filter models
No models match the current search and filters.
| Rank | Model | Company | No answer Count | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #257 | Mistral Small 4 medium | Mistral | 1 | 5.1 | $0.097 | 5/22 | 10.8s |
| #258 | Granite 4.2 8B low | IBM Granite | 1 | 5.1 | $0.012 | 7/22 | 54.4s |
| #259 | Qwen3 Coder Next none | Qwen | 1 | 5.1 | $0.026 | 5/22 | 8.63s |
| #260 | MiMo-V2.5 none | Xiaomi | 1 | 5.1 | $0.025 | 5/22 | 4.68s |
| #265 | Mercury 2.5 Preview low | Inception | 1 | 5.0 | $0.011 | 6/22 | 1.31s |
| #268 | GPT-4o-mini none | OpenAI | 1 | 5.0 | $0.010 | 5/22 | 1.92s |
| #276 | GPT-5.4 Nano none | OpenAI | 1 | 4.8 | $0.041 | 4/22 | 2.56s |
| #277 | Trinity Large Thinking high | Arcee AI | 1 | 4.8 | $0.592 | 5/22 | 75.9s |
| #283 | Grok 4.1 Fast medium | X AI | 1 | 4.7 | $0.069 | 9/19 | 23.8s |
| #284 | Laguna M.1 medium | Poolside | 1 | 4.7 | $0.033 | 9/19 | 14.7s |
| #287 | Granite 4.2 8B high | IBM Granite | 1 | 4.6 | $0.084 | 4/22 | 235.9s |
| #288 | Qwen3 Coder Next medium | Qwen | 1 | 4.6 | $0.034 | 3/22 | 9.07s |
| #290 | Laguna S 2.1 none | Poolside | 1 | 4.5 | $0.022 | 2/22 | 11.7s |
| #313 | Nemotron 3 Nano Omni 30b A3b Reasoning medium | NVIDIA | 1 | 3.4 | $0.000 | 4/19 | 17.1s |