Combined: No answer
Combined
No answer
See which AI models are most likely to hit No answer on Combined, so you can spot weak points faster. Sort by: Tests Correct ↑.
Failure Reasons
32/32
Filter models
No models match the current search and filters.
| Rank | Model | Company | No answer Count | Category Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #174 | MiMo-V2.5 none | Xiaomi | 1 | 3.0 | $0.025 | 0/2 | 28.9s |
| #178 | MiniMax M2.7 medium | Minimax | 1 | 3.8 | $0.163 | 0/2 | 72.1s |
| #180 | GPT-4o-mini none | OpenAI | 1 | 3.0 | $0.010 | 0/2 | 6.32s |
| #186 | GPT-5.4 Nano none | OpenAI | 1 | 3.0 | $0.041 | 0/2 | 14.7s |
| #192 | Laguna M.1 medium | Poolside | 1 | 1.5 | $0.033 | 0/1 | 53.1s |
| #193 | Qwen3 Coder Next medium | Qwen | 1 | 3.0 | $0.032 | 0/2 | 14.6s |
| #196 | MiniMax M2.5 medium | Minimax | 1 | 3.7 | $0.340 | 0/2 | 83.2s |
| #29 | GPT-5 Mini medium | OpenAI | 1 | 7.3 | $0.237 | 1/2 | 99.8s |
| #31 | Gemini 3.5 Flash-Lite high | 1 | 7.3 | $0.584 | 1/2 | 29.2s | |
| #34 | GPT-5.2 Chat none | OpenAI | 1 | 7.3 | $0.604 | 1/2 | 13.9s |
| #36 | Inkling medium | Thinkingmachines | 1 | 7.3 | $0.391 | 1/2 | 41.2s |
| #39 | Seed-2.0-Lite medium | Bytedance Seed | 1 | 6.4 | $0.234 | 1/2 | 58.5s |
| #77 | Grok 4.3 medium | X AI | 1 | 6.5 | $0.779 | 1/2 | 55.1s |
| #84 | Seed-2.0-Mini medium | Bytedance Seed | 1 | 7.3 | $0.101 | 1/2 | 282.3s |
| #95 | Gemini 3.5 Flash-Lite low | 1 | 6.3 | $0.145 | 1/2 | 8.96s |