Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

Qwen: Qwen3.5 Plus 2026-02-15 vs xAI: Grok 4.20

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-10

Kipimo Qwen3.5 Plus 2026-02-15 Qwen3.5 Plus 2026-02-15 none Toleo: 2026-02-15 Grok 4.20 Grok 4.20 medium Toleo: 2026-03-31
Alama 6.5 6.9
Nafasi #78 #68
Uaminifu 10.0 10.0
Uthabiti 9.3 8.3
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 50.9% 63.2%
Majaribio yasiyo thabiti 2 4
Jumla ya uendeshaji 57 57
Gharama kwa matokeo 0.183 7.559
Jumla ya gharama $0.017 $0.756
Bei ya ingizo $0.260 / 1M $1.250 / 1M
Bei ya toleo $1.560 / 1M $2.500 / 1M
Tokeni za matokeo 2,472 1,784
Tokeni za hoja 0 128,233
Muda wa majibu (wastani) 2.49s 14.53s
Muda wa majibu (upeo) 6.65s 63.48s
Muda wa majibu (jumla) 32.33s 276.06s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 4.8 10.0 25.0% 0 1.91s 517 0
Grok 4.20 8.2 7.9 83.3% 1 3.95s 287 8,312
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 6.3 3.7 33.3% 1 3.63s 443 0
Grok 4.20 4.3 1.1 66.7% 1 24.33s 250 12,804
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 3.0 10.0 0.0% 0 6.65s 314 0
Grok 4.20 10.0 10.0 100.0% 0 17.40s 232 9,556
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 10.0 10.0 100.0% 0 1.89s 243 0
Grok 4.20 10.0 10.0 100.0% 0 4.17s 180 5,333
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 5.3 10.0 33.3% 0 1.17s 17 0
Grok 4.20 5.3 10.0 33.3% 0 27.03s 375 49,339
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 4.4 3.0 33.3% 1 2.26s 117 0
Grok 4.20 3.9 2.6 33.3% 1 24.48s 65 6,440
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 10.0 10.0 100.0% 0 1.67s 72 0
Grok 4.20 7.3 6.0 83.3% 1 4.42s 40 5,474
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 7.7 10.0 66.7% 0 2.82s 516 0
Grok 4.20 7.7 10.0 66.7% 0 6.20s 149 7,913
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 10.0 10.0 100.0% 0 3.33s 222 0
Grok 4.20 3.0 10.0 0.0% 0 13.68s 197 6,620
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5 Plus 2026-02-15 3.0 10.0 0.0% 0 1.11s 11 0
Grok 4.20 3.0 10.0 0.0% 0 63.48s 9 16,442

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho