Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

xAI: Grok 4.20 Multi-Agent Beta vs Xiaomi: MiMo-V2-Omni

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-21

Kipimo Grok 4.20 Multi-Agent Beta Grok 4.20 Multi-Agent Beta medium Toleo: 2026-03-12 MiMo-V2-Omni MiMo-V2-Omni none Toleo: 2026-03-18
Alama 6.2 6.4
Nafasi #47 #43
Uthabiti 7.2 10.0
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 54.9% 47.1%
Majaribio yasiyo thabiti 6 0
Jumla ya uendeshaji 51 17
Gharama kwa matokeo 82.962 0.069
Jumla ya gharama $4.978 $0.006
Bei ya ingizo $2.000 / 1M $0.400 / 1M
Bei ya toleo $6.000 / 1M $2.000 / 1M
Tokeni za matokeo 298,948 469
Tokeni za hoja 296,529 0
Muda wa majibu (wastani) 8.64s 2.01s
Muda wa majibu (upeo) 35.28s 6.81s
Muda wa majibu (jumla) 129.64s 34.09s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 6.9 5.8 75.0% 2 3.46s 33,706 33,077
MiMo-V2-Omni 4.8 10.0 25.0% 0 1.10s 74 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
MiMo-V2-Omni 3.0 10.0 0.0% 0 2.47s 110 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 10.0 10.0 100.0% 0 5.54s 25,306 25,051
MiMo-V2-Omni 10.0 10.0 100.0% 0 1.69s 83 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 2.9 7.2 11.1% 1 24.67s 164,609 163,647
MiMo-V2-Omni 5.3 10.0 33.3% 0 1.14s 8 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 5.8 2.8 66.7% 1 6.40s 15,848 15,746
MiMo-V2-Omni 4.5 10.0 0.0% 0 1.19s 37 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 8.3 10.0 50.0% 0 4.63s 25,457 25,322
MiMo-V2-Omni 6.5 10.0 50.0% 0 4.18s 22 0
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 7.2 5.1 77.8% 2 5.01s 34,022 33,686
MiMo-V2-Omni 8.0 10.0 66.7% 0 2.71s 58 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi-Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
MiMo-V2-Omni 10.0 10.0 100.0% 0 2.76s 77 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho