Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

OpenAI: GPT-5 Nano vs Grok 4.20 Multi Agent Beta

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-04

Kipimo GPT-5 Nano GPT-5 Nano medium Toleo: 2025-08-07 Grok 4.20 Multi Agent Beta Grok 4.20 Multi Agent Beta medium Toleo: 2026-03-12
Alama 6.2 6.2
Nafasi #54 #55
Uthabiti 6.7 7.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 58.8% 54.9%
Majaribio yasiyo thabiti 7 6
Jumla ya uendeshaji 51 51
Gharama kwa matokeo 0.864 82.962
Jumla ya gharama $0.061 $4.978
Bei ya ingizo $0.050 / 1M $0.000 / 1M
Bei ya toleo $0.400 / 1M $0.000 / 1M
Tokeni za matokeo 4,500 298,948
Tokeni za hoja 143,296 296,529
Muda wa majibu (wastani) 44.47s 8.64s
Muda wa majibu (upeo) 204.02s 35.28s
Muda wa majibu (jumla) 444.74s 129.64s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 6.5 7.9 58.3% 1 25.50s 1,221 21,184
Grok 4.20 Multi Agent Beta 6.9 5.8 75.0% 2 3.46s 33,706 33,077
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 10.0 10.0 100.0% 0 65.96s 578 17,984
Grok 4.20 Multi Agent Beta 3.0 10.0 0.0% 0 0ms 0 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 3.7 1.7 50.0% 2 21.42s 453 10,560
Grok 4.20 Multi Agent Beta 10.0 10.0 100.0% 0 5.54s 25,306 25,051
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 5.2 4.4 55.6% 2 204.02s 237 64,448
Grok 4.20 Multi Agent Beta 2.9 7.2 11.1% 1 24.67s 164,609 163,647
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 4.1 10.0 0.0% 0 17.51s 202 4,608
Grok 4.20 Multi Agent Beta 5.8 2.8 66.7% 1 6.40s 15,848 15,746
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 8.5 6.8 83.3% 1 11.90s 382 4,096
Grok 4.20 Multi Agent Beta 8.3 10.0 50.0% 0 4.63s 25,457 25,322
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 5.3 7.2 44.4% 1 19.81s 869 13,440
Grok 4.20 Multi Agent Beta 7.2 5.1 77.8% 2 5.01s 34,022 33,686
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5 Nano 10.0 10.0 100.0% 0 33.30s 558 6,976
Grok 4.20 Multi Agent Beta 3.0 10.0 0.0% 0 0ms 0 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho