Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

xAI: Grok 4.20 Beta vs Xiaomi: MiMo-V2-Pro

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-20

Kipimo Grok 4.20 Beta Grok 4.20 Beta medium Toleo: 2026-03-12 MiMo-V2-Pro MiMo-V2-Pro medium Toleo: 2026-03-18
Alama 7.9 8.0
Nafasi #22 #20
Uthabiti 9.0 8.5
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 72.6% 76.5%
Majaribio yasiyo thabiti 2 3
Jumla ya uendeshaji 51 45
Gharama kwa matokeo 5.525 1.110
Jumla ya gharama $0.608 $0.123
Bei ya ingizo $2.000 / 1M $1.000 / 1M
Bei ya toleo $6.000 / 1M $3.000 / 1M
Tokeni za matokeo 1,487 1,875
Tokeni za hoja 87,922 26,959
Muda wa majibu (wastani) 8.54s 9.78s
Muda wa majibu (upeo) 24.21s 64.71s
Muda wa majibu (jumla) 145.26s 156.45s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 8.7 7.9 91.7% 1 3.16s 268 7,583
MiMo-V2-Pro 10.0 10.0 100.0% 0 3.06s 223 1,107
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 10.0 10.0 100.0% 0 20.93s 227 12,212
MiMo-V2-Pro 4.7 1.6 66.7% 1 64.71s 380 14,186
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 10.0 10.0 100.0% 0 4.01s 180 5,281
MiMo-V2-Pro 7.3 5.8 83.3% 1 17.20s 260 7,484
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 5.3 10.0 33.3% 0 21.33s 251 40,255
MiMo-V2-Pro 5.3 10.0 33.3% 0 6.00s 155 1,048
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 10.0 10.0 100.0% 0 5.78s 72 3,440
MiMo-V2-Pro 10.0 10.0 100.0% 0 4.06s 198 424
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 8.3 10.0 50.0% 0 4.97s 57 7,107
MiMo-V2-Pro 9.9 10.0 100.0% 0 3.36s 83 667
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 8.2 7.2 88.9% 1 3.85s 249 6,660
MiMo-V2-Pro 7.0 7.2 55.6% 1 4.71s 313 1,179
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Beta 3.0 10.0 0.0% 0 12.39s 183 5,384
MiMo-V2-Pro 10.0 10.0 100.0% 0 8.19s 263 864

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho