Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

xAI: Grok 4.20 vs Xiaomi: MiMo-V2-Flash

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-10

Kipimo Grok 4.20 Grok 4.20 medium Toleo: 2026-03-31 MiMo-V2-Flash MiMo-V2-Flash medium Toleo: 2025-12-16
Alama 6.9 7.2
Nafasi #68 #56
Uaminifu 10.0 10.0
Uthabiti 8.3 8.7
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 63.2% 66.7%
Majaribio yasiyo thabiti 4 3
Jumla ya uendeshaji 57 57
Gharama kwa matokeo 7.559 0.341
Jumla ya gharama $0.756 $0.038
Bei ya ingizo $1.250 / 1M $0.100 / 1M
Bei ya toleo $2.500 / 1M $0.300 / 1M
Tokeni za matokeo 1,784 12,399
Tokeni za hoja 128,233 115,182
Muda wa majibu (wastani) 14.53s 21.71s
Muda wa majibu (upeo) 63.48s 96.01s
Muda wa majibu (jumla) 276.06s 282.29s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 8.2 7.9 83.3% 1 3.95s 287 8,312
MiMo-V2-Flash 8.1 7.9 83.3% 1 15.85s 1,674 23,559
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 4.3 1.1 66.7% 1 24.33s 250 12,804
MiMo-V2-Flash 4.7 1.6 66.7% 1 13.03s 428 3,648
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 10.0 10.0 100.0% 0 17.40s 232 9,556
MiMo-V2-Flash 9.8 10.0 100.0% 0 75.68s 442 26,859
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 10.0 10.0 100.0% 0 4.17s 180 5,333
MiMo-V2-Flash 6.5 10.0 50.0% 0 0ms 153 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 5.3 10.0 33.3% 0 27.03s 375 49,339
MiMo-V2-Flash 5.9 7.2 55.6% 1 96.01s 8,374 42,461
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 3.9 2.6 33.3% 1 24.48s 65 6,440
MiMo-V2-Flash 4.0 10.0 0.0% 0 4.20s 87 488
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 7.3 6.0 83.3% 1 4.42s 40 5,474
MiMo-V2-Flash 10.0 10.0 100.0% 0 4.28s 75 3,504
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 7.7 10.0 66.7% 0 6.20s 149 7,913
MiMo-V2-Flash 7.7 10.0 66.7% 0 3.77s 833 1,948
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 3.0 10.0 0.0% 0 13.68s 197 6,620
MiMo-V2-Flash 10.0 10.0 100.0% 0 27.78s 321 12,715
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 3.0 10.0 0.0% 0 63.48s 9 16,442
MiMo-V2-Flash 3.0 10.0 0.0% 0 1.96s 12 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho