AI BENCHY
Your ad here

Kategoria ya AI BENCHY

Orodha ya Utatuzi wa mafumbo

Ona ni modeli gani za AI zinafanya vizuri zaidi katika Utatuzi wa mafumbo, zipi zinabaki thabiti, na pengo kubwa liko wapi. Panga kwa: Kipimo ↑.

Modeli zilizoonyeshwa

15

Wastani wa Alama ya Utatuzi wa mafumbo

6.4

Nafasi Modeli Kampuni Alama ya Utatuzi wa mafumbo Alama Majaribio sahihi Muda wa majibu (wastani)
#25 Grok 4.20 Beta medium X AI 8.2 8.0 2/3 3.85s
#27 DeepSeek V3.2 medium DeepSeek 8.2 8.0 2/3 36.9s
#33 GLM 5.1 medium Z.ai 8.2 7.8 2/3 23.8s
#39 Seed-2.0-Mini medium Bytedance Seed 8.2 7.5 2/3 25.9s
#64 DeepSeek V3.2 none DeepSeek 8.5 6.1 2/3 7.37s
#14 Gemma 4 31B medium Google 8.8 8.3 2/3 27.6s
#6 Seed-2.0-Lite medium Bytedance Seed 9.0 8.6 2/3 11.0s
#7 GPT-5.3-Codex medium OpenAI 9.0 8.6 2/3 5.12s
#1 Gemini 3 Flash Preview medium Google 10.0 10.0 3/3 4.43s
#2 Gemini 3.1 Pro Preview medium Google 10.0 9.6 3/3 7.15s
#3 Claude Opus 4.7 medium Anthropic 10.0 9.2 3/3 2.51s
#4 Claude Opus 4.7 none Anthropic 10.0 9.2 3/3 2.58s
#5 Gemini 3 Flash Preview low Google 10.0 8.8 3/3 6.11s
#8 Qwen3.5 Plus 2026-02-15 medium Qwen 10.0 8.5 3/3 34.6s
#9 Qwen3.6 Plus Preview medium Qwen 10.0 8.5 3/3 6.11s

Modeli bora kwa Alama ya Utatuzi wa mafumbo

Alama ya Utatuzi wa mafumbo dhidi ya jumla ya gharama

Modeli bora kwa Muda wa majibu (wastani)