Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

OpenAI: GPT-5.4 vs Qwen: Qwen3.6 Flash

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-27

Kipimo GPT-5.4 GPT-5.4 medium Toleo: 2026-03-05 Qwen3.6 Flash Qwen3.6 Flash medium Toleo: 2026-04-20
Alama 8.2 8.1
Nafasi #21 #24
Uaminifu Haipo 10.0
Uthabiti 8.7 8.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 79.6% 79.6%
Majaribio yasiyo thabiti 3 4
Jumla ya uendeshaji 54 54
Gharama kwa matokeo 6.399 1.449
Jumla ya gharama $0.832 $0.174
Bei ya ingizo $2.500 / 1M $0.250 / 1M
Bei ya toleo $15.000 / 1M $1.500 / 1M
Tokeni za matokeo 2,169 2,804
Tokeni za hoja 48,732 107,210
Muda wa majibu (wastani) 18.63s 9.90s
Muda wa majibu (upeo) 100.41s 26.85s
Muda wa majibu (jumla) 335.26s 178.26s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 8.3 10.0 75.0% 0 4.11s 240 1,511
Qwen3.6 Flash 10.0 10.0 100.0% 0 6.10s 624 14,024
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 10.0 10.0 100.0% 0 13.03s 389 2,045
Qwen3.6 Flash 6.7 3.5 66.7% 1 25.84s 435 17,044
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 10.0 10.0 100.0% 0 20.57s 301 3,543
Qwen3.6 Flash 10.0 10.0 100.0% 0 20.28s 483 13,839
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 10.0 10.0 100.0% 0 5.32s 234 804
Qwen3.6 Flash 10.0 10.0 100.0% 0 9.65s 270 13,155
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 5.3 7.2 44.4% 1 74.27s 61 34,748
Qwen3.6 Flash 3.5 4.4 33.3% 2 14.65s 60 24,409
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 4.7 3.1 33.3% 1 4.92s 145 321
Qwen3.6 Flash 4.8 9.9 0.0% 0 9.88s 140 5,445
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 10.0 10.0 100.0% 0 3.11s 93 897
Qwen3.6 Flash 10.0 10.0 100.0% 0 6.05s 102 7,423
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 8.2 7.2 88.9% 1 9.13s 442 3,832
Qwen3.6 Flash 8.2 7.2 88.9% 1 6.17s 355 10,683
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
GPT-5.4 10.0 10.0 100.0% 0 13.28s 264 1,031
Qwen3.6 Flash 10.0 10.0 100.0% 0 4.00s 335 1,188

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho