Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

Qwen3.6 Plus Preview vs xAI: Grok 4.20

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-29

Kipimo Qwen3.6 Plus Preview Qwen3.6 Plus Preview medium Toleo: 2026-04-20 Inapatikana bure Grok 4.20 Grok 4.20 medium Toleo: 2026-03-31
Alama 8.5 7.0
Nafasi #14 #62
Uaminifu Haipo Haipo
Uthabiti 10.0 7.8
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 76.5% 66.7%
Majaribio yasiyo thabiti 0 5
Jumla ya uendeshaji 49 54
Gharama kwa matokeo 0.000 8.252
Jumla ya gharama $0.000 $0.743
Bei ya ingizo $0.000 / 1M $2.000 / 1M
Bei ya toleo $0.000 / 1M $6.000 / 1M
Tokeni za matokeo 1,756 1,744
Tokeni za hoja 77,213 109,882
Muda wa majibu (wastani) 13.94s 10.33s
Muda wa majibu (upeo) 43.55s 29.87s
Muda wa majibu (jumla) 237.01s 185.87s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 9.90s 207 7,557
Grok 4.20 8.2 7.9 83.3% 1 3.36s 280 8,476
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 34.95s 452 13,073
Grok 4.20 10.0 10.0 100.0% 0 17.40s 232 9,556
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 14.95s 270 10,706
Grok 4.20 10.0 10.0 100.0% 0 4.17s 180 5,333
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 3.0 10.0 0.0% 0 22.08s 49 26,895
Grok 4.20 5.3 10.0 33.3% 0 27.03s 375 49,339
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 5.1 10.0 0.0% 0 27.05s 111 5,232
Grok 4.20 5.8 2.8 66.7% 1 7.09s 47 4,252
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 7.54s 102 5,552
Grok 4.20 7.3 5.9 83.3% 1 4.42s 40 5,474
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 6.11s 298 6,868
Grok 4.20 6.4 7.7 55.6% 1 3.89s 143 8,028
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 5.87s 267 1,330
Grok 4.20 3.0 10.0 0.0% 0 13.68s 197 6,620
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview - - - - - - - -
Grok 4.20 4.3 1.1 66.7% 1 24.33s 250 12,804

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho