Kategori AI BENCHY
Peringkat Kecerdasan umum
Lihat model AI mana yang paling baik di Kecerdasan umum, mana yang tetap andal, dan di mana kesenjangan terbesar muncul. Urutkan berdasarkan: Tes benar ↓.
Model yang ditampilkan
15
Rata-rata Skor Kecerdasan umum
5.9
Model terbaik
Gemini 3 Flash Preview 10.0| Peringkat | Model | Perusahaan | Skor Kecerdasan umum | Skor | Tes benar | Waktu respons (rata-rata) |
|---|---|---|---|---|---|---|
| #101 | Mimo V2 Omni none | Xiaomi | 4.1 | 6.0 | 0/1 | 2.33s |
| #102 | Gemma 4 26B A4B none | 4.0 | 6.0 | 0/1 | 3.54s | |
| #103 | DeepSeek V4 Pro high | DeepSeek | 6.1 | 6.0 | 0/1 | 25.1s |
| #104 | Nemotron 3 Ultra 550b A55b none | NVIDIA | 5.0 | 6.0 | 0/1 | 13.5s |
| #105 | Nemotron 3 Super medium | NVIDIA | 4.1 | 5.8 | 0/1 | 6.91s |
| #106 | Grok 4.20 Beta none | X AI | 5.0 | 5.8 | 0/1 | 541ms |
| #107 | Laguna Xs.2 medium | Poolside | 3.0 | 5.8 | 0/1 | 0ms |
| #109 | GLM 5V Turbo none | Z.ai | 4.6 | 5.8 | 0/1 | 2.22s |
| #111 | Owl Alpha medium | Openrouter | 4.3 | 5.7 | 0/1 | 58.6s |
| #112 | GLM 5.1 none | Z.ai | 5.0 | 5.7 | 0/1 | 790ms |
| #113 | DeepSeek V4 Pro none | DeepSeek | 4.3 | 5.7 | 0/1 | 3.75s |
| #114 | Qwen3.5 Plus 2026-04-20 none | Qwen | 4.8 | 5.7 | 0/1 | 1.41s |
| #115 | Qwen3.5-27B none | Qwen | 5.0 | 5.7 | 0/1 | 2.51s |
| #116 | Hunter Alpha none | OpenRouter | 6.1 | 5.7 | 0/1 | 2.71s |
| #117 | Qwen3.5-35B-A3B none | Qwen | 6.5 | 5.6 | 0/1 | 1.19s |