Kategoria ya AI BENCHY
Orodha ya Akili ya jumla
Ona ni modeli gani za AI zinafanya vizuri zaidi katika Akili ya jumla, zipi zinabaki thabiti, na pengo kubwa liko wapi. Panga kwa: Kipimo ↑.
| Nafasi | Modeli | Kampuni | Alama ya Akili ya jumla | Alama | Majaribio sahihi | Muda wa majibu (wastani) |
|---|---|---|---|---|---|---|
| #4 | Claude Opus 4.7 none | Anthropic | 10.0 | 9.2 | 1/1 | 3.47s |
| #5 | Gemini 3 Flash Preview low | 10.0 | 8.8 | 1/1 | 3.68s | |
| #11 | Gemini 3.1 Flash Lite Preview high | 10.0 | 8.4 | 1/1 | 5.25s | |
| #12 | Gemini 3 PRO Preview medium | 10.0 | 8.4 | 1/1 | 9.34s | |
| #14 | Gemma 4 31B medium | 10.0 | 8.3 | 1/1 | 9.57s | |
| #17 | Gemini 3.1 Flash Lite Preview medium | 10.0 | 8.2 | 1/1 | 3.16s | |
| #21 | Gemini 3 Flash Preview none | 10.0 | 8.1 | 1/1 | 1.13s | |
| #23 | MiMo-V2-Pro medium | Xiaomi | 10.0 | 8.1 | 1/1 | 4.06s |
| #24 | Gemma 4 26B A4B medium | 10.0 | 8.0 | 1/1 | 29.8s | |
| #25 | Grok 4.20 Beta medium | X AI | 10.0 | 8.0 | 1/1 | 5.78s |
| #26 | Claude Sonnet 4.6 medium | Anthropic | 10.0 | 8.0 | 1/1 | 4.94s |
| #31 | GLM 5V Turbo medium | Z.ai | 10.0 | 7.8 | 1/1 | 11.1s |
| #33 | GLM 5.1 medium | Z.ai | 10.0 | 7.8 | 1/1 | 20.9s |
| #34 | Kimi K2.6 medium | Moonshot AI | 10.0 | 7.7 | 1/1 | 17.8s |
| #35 | MiMo-V2-Omni medium | Xiaomi | 10.0 | 7.7 | 1/1 | 2.86s |