Kategori AI BENCHY
Peringkat Kecerdasan umum
Lihat model AI mana yang paling baik di Kecerdasan umum, mana yang tetap andal, dan di mana kesenjangan terbesar muncul.
Model yang ditampilkan
13
Rata-rata Skor Kecerdasan umum
5.9
Model terbaik
Gemini 3 Flash Preview 10.0| Peringkat | Model | Perusahaan | Skor Kecerdasan umum | Skor | Tes benar | Waktu respons (rata-rata) |
|---|---|---|---|---|---|---|
| #29 | Qwen3.5-122B-A10B medium | Qwen | 3.4 | 7.8 | 0/1 | 34.1s |
| #72 | DeepSeek V3.2 medium | DeepSeek | 3.4 | 7.0 | 0/1 | 58.3s |
| #82 | Hy3 preview high | Tencent | 3.0 | 6.6 | 0/1 | 0ms |
| #89 | Hy3 preview low | Tencent | 3.0 | 6.4 | 0/1 | 0ms |
| #92 | Laguna M.1 medium | Poolside | 3.0 | 6.4 | 0/1 | 0ms |
| #93 | Qwen3.6 Plus Preview medium | Qwen | 3.0 | 6.3 | 0/1 | 0ms |
| #107 | Laguna Xs.2 medium | Poolside | 3.0 | 5.8 | 0/1 | 0ms |
| #145 | Laguna M.1 none | Poolside | 3.0 | 4.8 | 0/1 | 0ms |
| #146 | Laguna Xs.2 none | Poolside | 3.0 | 4.8 | 0/1 | 0ms |
| #149 | Nemotron 3 Nano Omni 30b A3b Reasoning medium | NVIDIA | 3.0 | 4.6 | 0/1 | 0ms |
| #162 | Nemotron 3 Nano Omni 30b A3b Reasoning none | NVIDIA | 3.0 | 4.1 | 0/1 | 0ms |
| #66 | Qwen3.5-35B-A3B medium | Qwen | 2.8 | 7.1 | 0/1 | 30.3s |
| #161 | Qwen3.5-9B medium | Qwen | 2.8 | 4.2 | 0/1 | 226.4s |