Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

Grok 4.20 Multi Agent Beta vs Z.ai: GLM 5V Turbo

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-02

Kipimo Grok 4.20 Multi Agent Beta Grok 4.20 Multi Agent Beta medium Toleo: 2026-03-12 GLM 5V Turbo GLM 5V Turbo none Toleo: 2026-04-01
Alama 6.2 6.0
Nafasi #53 #55
Uthabiti 7.2 10.0
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 54.9% 41.2%
Majaribio yasiyo thabiti 6 0
Jumla ya uendeshaji 51 51
Gharama kwa matokeo 82.962 0.588
Jumla ya gharama $4.978 $0.042
Bei ya ingizo $0.000 / 1M $1.200 / 1M
Bei ya toleo $0.000 / 1M $4.000 / 1M
Tokeni za matokeo 298,948 1,388
Tokeni za hoja 296,529 0
Muda wa majibu (wastani) 8.64s 2.97s
Muda wa majibu (upeo) 35.28s 6.51s
Muda wa majibu (jumla) 129.64s 50.57s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 6.9 5.8 75.0% 2 3.46s 33,706 33,077
GLM 5V Turbo 4.8 10.0 25.0% 0 3.13s 281 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
GLM 5V Turbo 3.0 10.0 0.0% 0 6.51s 276 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 10.0 10.0 100.0% 0 5.54s 25,306 25,051
GLM 5V Turbo 10.0 10.0 100.0% 0 3.81s 204 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 2.9 7.2 11.1% 1 24.67s 164,609 163,647
GLM 5V Turbo 5.3 10.0 33.3% 0 2.09s 24 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 5.8 2.8 66.7% 1 6.40s 15,848 15,746
GLM 5V Turbo 4.6 10.0 0.0% 0 2.22s 114 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 8.3 10.0 50.0% 0 4.63s 25,457 25,322
GLM 5V Turbo 6.5 10.0 50.0% 0 1.97s 60 0
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 7.2 5.1 77.8% 2 5.01s 34,022 33,686
GLM 5V Turbo 5.3 10.0 33.3% 0 2.22s 207 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
GLM 5V Turbo 10.0 10.0 100.0% 0 4.86s 222 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho