AI BENCHY Category Failures
Anti-AI Tricks: Did not follow instructions
Anti-AI Tricks
Did not follow instructions
See which AI models are most likely to hit Did not follow instructions on Anti-AI Tricks, so you can spot weak points faster. Sort by: Tests Correct ↓.
Failure Reasons
| Rank | Model | Company | Did not follow instructions Count | Category Score | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|
| #92 | Qwen3 Coder Next medium | Qwen | 1 | 3.5 | 0/4 | 8.64s |
| #95 | Grok 4.1 Fast none | X AI | 1 | 3.2 | 0/4 | 1.07s |