Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Qwen3.6 Plus Preview vs Grok 4.20 Beta

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-29

Kipimo Qwen3.6 Plus Preview Qwen3.6 Plus Preview medium Toleo: 2026-04-20 Inapatikana bure Grok 4.20 Beta Grok 4.20 Beta none Toleo: 2026-03-12
Alama 8.5 5.3
Nafasi #14 #104
Uaminifu Haipo Haipo
Uthabiti 10.0 9.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 76.5% 29.6%
Majaribio yasiyo thabiti 0 2
Jumla ya uendeshaji 49 52
Gharama kwa matokeo 0.000 2.255
Jumla ya gharama $0.000 $0.091
Bei ya ingizo $0.000 / 1M $0.000 / 1M
Bei ya toleo $0.000 / 1M $0.000 / 1M
Tokeni za matokeo 1,756 1,591
Tokeni za hoja 77,213 0
Muda wa majibu (wastani) 13.94s 1.19s
Muda wa majibu (upeo) 43.55s 6.48s
Muda wa majibu (jumla) 237.01s 21.37s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 9.90s 207 7,557
Grok 4.20 Beta 4.0 8.4 16.7% 1 597ms 251 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 34.95s 452 13,073
Grok 4.20 Beta 3.0 10.0 0.0% 0 6.48s 282 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 14.95s 270 10,706
Grok 4.20 Beta 10.0 10.0 100.0% 0 601ms 197 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 3.0 10.0 0.0% 0 22.08s 49 26,895
Grok 4.20 Beta 3.0 10.0 0.0% 0 611ms 160 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 5.1 10.0 0.0% 0 27.05s 111 5,232
Grok 4.20 Beta 5.0 10.0 0.0% 0 541ms 87 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 7.54s 102 5,552
Grok 4.20 Beta 4.8 10.0 0.0% 0 687ms 60 0
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 6.11s 298 6,868
Grok 4.20 Beta 5.9 7.2 55.6% 1 541ms 291 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview 10.0 10.0 100.0% 0 5.87s 267 1,330
Grok 4.20 Beta 10.0 10.0 100.0% 0 4.79s 189 0
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.6 Plus Preview - - - - - - - -
Grok 4.20 Beta 5.5 10.0 0.0% 0 1.14s 74 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho