Kategori AI BENCHY
Peringkat Trik anti-AI
Lihat model AI mana yang paling baik di Trik anti-AI, mana yang tetap andal, dan di mana kesenjangan terbesar muncul. Urutkan berdasarkan: Tes benar ↓.
| Peringkat | Model | Perusahaan | Skor Trik anti-AI | Skor | Tes benar | Waktu respons (rata-rata) |
|---|---|---|---|---|---|---|
| #116 | Hunter Alpha none | OpenRouter | 3.5 | 5.7 | 0/4 | 3.81s |
| #117 | Qwen3.5-35B-A3B none | Qwen | 3.4 | 5.6 | 0/4 | 1.43s |
| #118 | Qwen3.6 27B none | Qwen | 3.8 | 5.6 | 0/4 | 2.83s |
| #120 | Mimo V2 PRO none | Xiaomi | 3.5 | 5.6 | 0/4 | 1.80s |
| #121 | Owl Alpha none | Openrouter | 3.4 | 5.5 | 0/4 | 2.78s |
| #123 | MiMo-V2.5-Pro none | Xiaomi | 3.3 | 5.5 | 0/4 | 2.67s |
| #125 | GPT-5.4 none | OpenAI | 3.2 | 5.5 | 0/4 | 1.21s |
| #128 | Qwen3.6 Flash none | Qwen | 3.1 | 5.4 | 0/4 | 1.63s |
| #133 | DeepSeek V3.2 none | DeepSeek | 3.2 | 5.2 | 0/4 | 9.35s |
| #134 | GLM 5 Turbo none | Z.ai | 3.0 | 5.2 | 0/4 | 2.84s |
| #135 | Kimi K2.5 none | Moonshot AI | 3.6 | 5.2 | 0/4 | 6.24s |
| #139 | DeepSeek V4 Flash none | DeepSeek | 3.0 | 5.0 | 0/4 | 20.2s |
| #140 | Qwen3 Coder Next none | Qwen | 3.6 | 4.9 | 0/4 | 3.31s |
| #142 | Mistral Small 4 none | Mistral | 3.4 | 4.9 | 0/4 | 395ms |
| #143 | MiMo-V2.5 none | Xiaomi | 3.5 | 4.9 | 0/4 | 2.19s |