Urambazaji
AI BENCHY
Linganisha Chati Mbinu
❤️ Made by XCS
Your ad here

AI BENCHY Compare

OpenAI: GPT-5.4 vs Qwen: Qwen3.5 Plus 2026-02-15

Linganisha:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-06

Kipimo OpenAI: GPT-5.4 medium Toleo: 2026-03-05 Qwen: Qwen3.5 Plus 2026-02-15 none Toleo: 2026-02-15
Nafasi #9 #29
Wastani wa alama 8.0 6.2
Uthabiti 8.5 9.6
Gharama kwa matokeo 6.601 0.172
Jumla ya gharama $0.793 $0.016
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 83.3% 58.3%
Majaribio yasiyo thabiti 3 1
Jumla ya uendeshaji 48 (16 x 3) 48 (16 x 3)
Tokeni za matokeo 1,756 2,015
Tokeni za hoja 46,642 0
Muda wa majibu (wastani) 20.05s 2.65s
Muda wa majibu (upeo) 100.41s 6.65s
Muda wa majibu (jumla) 320.87s 26.52s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Wastani wa alama vs Muda wa majibu (wastani)

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 10.0 10.0 100.0% 0 5.02s 216 1,466
Qwen: Qwen3.5 Plus 2026-02-15 4.0 10.0 33.3% 0 2.74s 514 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 10.0 10.0 100.0% 0 20.57s 301 3,543
Qwen: Qwen3.5 Plus 2026-02-15 10.0 10.0 0.0% 0 6.65s 314 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 9.9 10.0 100.0% 0 5.32s 234 804
Qwen: Qwen3.5 Plus 2026-02-15 9.9 10.0 100.0% 0 1.89s 243 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 4.0 7.2 44.4% 1 74.27s 61 34,748
Qwen: Qwen3.5 Plus 2026-02-15 4.0 10.0 33.3% 0 1.17s 17 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 5.0 3.1 33.3% 1 4.92s 145 321
Qwen: Qwen3.5 Plus 2026-02-15 4.0 3.0 33.3% 1 2.26s 117 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 10.0 10.0 100.0% 0 3.11s 93 897
Qwen: Qwen3.5 Plus 2026-02-15 10.0 10.0 100.0% 0 1.67s 72 0
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 7.0 7.2 88.9% 1 9.13s 442 3,832
Qwen: Qwen3.5 Plus 2026-02-15 7.0 10.0 66.7% 0 2.82s 516 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.4 10.0 10.0 100.0% 0 13.28s 264 1,031
Qwen: Qwen3.5 Plus 2026-02-15 10.0 10.0 100.0% 0 3.33s 222 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho