Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

Grok 4.20 Multi Agent Beta vs GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-02

Kipimo Grok 4.20 Multi Agent Beta Grok 4.20 Multi Agent Beta medium Toleo: 2026-03-12 GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT none Toleo: Tarehe ya kutolewa haijulikani
Alama 6.2 3.0
Nafasi #53 #88
Uthabiti 7.2 10.0
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 54.9% 0.0%
Majaribio yasiyo thabiti 6 0
Jumla ya uendeshaji 51 48
Gharama kwa matokeo 82.962 0.000
Jumla ya gharama $4.978 $0.000
Bei ya ingizo $0.000 / 1M $0.000 / 1M
Bei ya toleo $0.000 / 1M $0.000 / 1M
Tokeni za matokeo 298,948 0
Tokeni za hoja 296,529 0
Muda wa majibu (wastani) 8.64s 0ms
Muda wa majibu (upeo) 35.28s 0ms
Muda wa majibu (jumla) 129.64s 0ms

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 6.9 5.8 75.0% 2 3.46s 33,706 33,077
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 10.0 10.0 100.0% 0 5.54s 25,306 25,051
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 2.9 7.2 11.1% 1 24.67s 164,609 163,647
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 5.8 2.8 66.7% 1 6.40s 15,848 15,746
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 8.3 10.0 50.0% 0 4.63s 25,457 25,322
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 7.2 5.1 77.8% 2 5.01s 34,022 33,686
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Grok 4.20 Multi Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
GLM 5v Turbo X Ai/grok 4.20 Google/gemma 4 31b IT 3.0 10.0 0.0% 0 0ms 0 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho