Combined: No answer
Combined
No answer
See which AI models are most likely to hit No answer on Combined, so you can spot weak points faster. Sort by: Response Time (avg) ↓.
Failure Reasons
32/32
Filter models
No models match the current search and filters.
| Rank | Model | Company | No answer Count | Category Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #125 | Qwen3.5-35B-A3B medium | Qwen | 1 | 3.8 | $0.837 | 0/2 | 512.8s |
| #84 | Seed-2.0-Mini medium | Bytedance Seed | 1 | 7.3 | $0.101 | 1/2 | 282.3s |
| #134 | GPT-5 Nano medium | OpenAI | 1 | 6.4 | $0.114 | 1/2 | 146.9s |
| #29 | GPT-5 Mini medium | OpenAI | 1 | 7.3 | $0.237 | 1/2 | 99.8s |
| #196 | MiniMax M2.5 medium | Minimax | 1 | 3.7 | $0.340 | 0/2 | 83.2s |
| #144 | Kimi K2.6 none | Moonshot AI | 1 | 3.0 | $0.184 | 0/2 | 77.8s |
| #178 | MiniMax M2.7 medium | Minimax | 1 | 3.8 | $0.163 | 0/2 | 72.1s |
| #161 | Kimi K2.5 none | Moonshot AI | 1 | 2.8 | $0.127 | 0/2 | 61.0s |
| #39 | Seed-2.0-Lite medium | Bytedance Seed | 1 | 6.4 | $0.234 | 1/2 | 58.5s |
| #77 | Grok 4.3 medium | X AI | 1 | 6.5 | $0.779 | 1/2 | 55.1s |
| #192 | Laguna M.1 medium | Poolside | 1 | 1.5 | $0.033 | 0/1 | 53.1s |
| #157 | GLM 5.1 none | Z.ai | 1 | 2.8 | $0.164 | 0/2 | 46.9s |
| #36 | Inkling medium | Thinkingmachines | 1 | 7.3 | $0.391 | 1/2 | 41.2s |
| #167 | Qwen3.6 35B A3B none | Qwen | 1 | 3.8 | $0.061 | 0/2 | 39.5s |
| #173 | Mistral Small 4 medium | Mistral | 1 | 3.0 | $0.096 | 0/2 | 32.4s |