AI BENCHY
Advertise here

Kategoria ya AI BENCHY

Orodha ya Mahususi kwa domeni

Ona ni modeli gani za AI zinafanya vizuri zaidi katika Mahususi kwa domeni, zipi zinabaki thabiti, na pengo kubwa liko wapi. Panga kwa: Muda wa majibu (wastani) ↓.

Modeli zilizoonyeshwa

15

Wastani wa Alama ya Mahususi kwa domeni

4.8

Modeli bora

MiniMax M2.5 2.9
Nafasi Modeli Kampuni Alama ya Mahususi kwa domeni Alama Majaribio sahihi Muda wa majibu (wastani)
#129 MiniMax M2.5 medium Minimax 2.9 5.3 0/3 237.3s
#67 MiniMax M3 medium Minimax 5.5 7.1 1/3 233.1s
#103 DeepSeek V4 Pro high DeepSeek 2.9 6.0 0/3 205.7s
#94 GPT-5 Nano medium OpenAI 5.2 6.3 1/3 204.0s
#60 Kimi K2.6 medium Moonshot AI 5.3 7.2 1/3 202.4s
#38 Grok 4.3 medium X AI 5.3 7.6 1/3 181.7s
#158 GLM 4.7 Flash medium Z.ai 3.5 4.4 0/3 174.6s
#62 Step 3.5 Flash medium Stepfun 5.3 7.2 1/3 170.5s
#9 GPT-5.5 medium OpenAI 5.3 8.8 1/3 164.1s
#47 Grok Build 0.1 medium X AI 5.3 7.4 1/3 158.0s
#71 Step 3.7 Flash high Stepfun 4.1 7.0 0/3 149.6s
#49 Qwen3.5-Flash medium Qwen 5.3 7.4 1/3 146.5s
#53 Gemini 3.1 Flash Lite high Google 3.6 7.3 0/3 139.9s
#161 Qwen3.5-9B medium Qwen 3.6 4.2 0/3 137.7s
#76 Kimi K2.5 medium Moonshot AI 3.5 6.8 0/3 137.3s

Modeli bora kwa Alama ya Mahususi kwa domeni

Alama ya Mahususi kwa domeni dhidi ya jumla ya gharama

Modeli bora kwa Muda wa majibu (wastani)