AI BENCHY
Advertise here

Kategoria ya AI BENCHY

Orodha ya Utatuzi wa mafumbo

Ona ni modeli gani za AI zinafanya vizuri zaidi katika Utatuzi wa mafumbo, zipi zinabaki thabiti, na pengo kubwa liko wapi. Panga kwa: Majaribio sahihi ↑.

Modeli zilizoonyeshwa

15

Wastani wa Alama ya Utatuzi wa mafumbo

6.7

Modeli bora

GPT-5.4 Nano 4.1
Nafasi Modeli Kampuni Alama ya Utatuzi wa mafumbo Alama Majaribio sahihi Muda wa majibu (wastani)
#79 Hunter Alpha medium OpenRouter 6.1 6.7 1/3 5.35s
#80 Mimo V2 Omni medium Xiaomi 5.9 6.7 1/3 2.38s
#81 Mercury 2 medium Inception 5.4 6.6 1/3 949ms
#84 Grok 4.20 Multi Agent Beta medium X AI 6.7 6.6 1/3 5.19s
#85 Gemma 4 31B none Google 6.5 6.5 1/3 4.23s
#86 Grok 4.1 Fast medium X AI 5.3 6.5 1/3 7.40s
#87 Gemini 3.1 Flash Lite minimal Google 6.0 6.4 1/3 2.15s
#89 Hy3 preview low Tencent 5.3 6.4 1/3 7.51s
#90 Gemini 3.1 Flash Lite none Google 6.3 6.4 1/3 720ms
#92 Laguna M.1 medium Poolside 5.3 6.4 1/3 10.2s
#93 Qwen3.6 Plus Preview medium Qwen 5.3 6.3 1/3 7.52s
#94 GPT-5 Nano medium OpenAI 5.3 6.3 1/3 20.6s
#99 gpt-oss-120b medium OpenAI 5.3 6.1 1/3 21.7s
#100 Grok Build 0.1 none X AI 6.4 6.0 1/3 9.55s
#102 Gemma 4 26B A4B none Google 6.2 6.0 1/3 744ms

Modeli bora kwa Alama ya Utatuzi wa mafumbo

Alama ya Utatuzi wa mafumbo dhidi ya jumla ya gharama

Modeli bora kwa Muda wa majibu (wastani)