Kategoria ya AI BENCHY
Orodha ya Mwito wa zana
Ona ni modeli gani za AI zinafanya vizuri zaidi katika Mwito wa zana, zipi zinabaki thabiti, na pengo kubwa liko wapi. Panga kwa: Muda wa majibu (wastani) ↑.
| Nafasi | Modeli | Kampuni | Alama ya Mwito wa zana | Alama | Majaribio sahihi | Muda wa majibu (wastani) |
|---|---|---|---|---|---|---|
| #86 | Grok 4.1 Fast medium | X AI | 2.8 | 6.5 | 0/1 | 27.7s |
| #64 | MiMo-V2-Flash medium | Xiaomi | 10.0 | 7.2 | 1/1 | 27.8s |
| #76 | Kimi K2.5 medium | Moonshot AI | 10.0 | 6.8 | 1/1 | 31.7s |
| #94 | GPT-5 Nano medium | OpenAI | 10.0 | 6.3 | 1/1 | 33.3s |
| #156 | Hy3 preview none | Tencent | 10.0 | 4.4 | 1/1 | 33.8s |
| #72 | DeepSeek V3.2 medium | DeepSeek | 10.0 | 7.0 | 1/1 | 34.8s |
| #105 | Nemotron 3 Super medium | NVIDIA | 10.0 | 5.8 | 1/1 | 39.7s |
| #102 | Gemma 4 26B A4B none | 10.0 | 6.0 | 1/1 | 57.1s | |
| #31 | DeepSeek V4 Flash high | DeepSeek | 10.0 | 7.7 | 1/1 | 74.7s |
| #139 | DeepSeek V4 Flash none | DeepSeek | 10.0 | 5.0 | 1/1 | 77.9s |
| #82 | Hy3 preview high | Tencent | 10.0 | 6.6 | 1/1 | 78.8s |
| #73 | Seed-2.0-Mini medium | Bytedance Seed | 10.0 | 6.9 | 1/1 | 88.7s |
| #75 | Ring-2.6-1T medium | Inclusionai | 10.0 | 6.9 | 1/1 | 104.4s |