Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

OpenAI: GPT-5.2 vs Qwen: Qwen3.5-27B

Jina la modeli:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe : 2026-02-27 15:16

Muhtasari

Kipimo OpenAI: GPT-5.2 medium Toleo: Tarehe ya kutolewa haijulikani Qwen: Qwen3.5-27B none Toleo: Tarehe ya kutolewa haijulikani
Nafasi #12 #29
Alama 6.93 4.70
Uthabiti 8.22 9.93
Gharama kwa matokeo 2.780 0.190
Jumla ya gharama $0.251 $0.010
Majaribio sahihi
Majaribio yenye makosa 5 9
Kiwango cha kupita kwa kila jaribio 76.2% 35.7%
Majaribio yasiyo thabiti 3 0
Tokeni za matokeo 1,869 1,600
Tokeni za hoja 14,190 0

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.2 7.00 7.28 77.8% 1 549 2,002
Qwen: Qwen3.5-27B 4.00 10.00 33.3% 0 264 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.2 10.00 10.00 100.0% 0 234 499
Qwen: Qwen3.5-27B 9.88 10.00 100.0% 0 243 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.2 4.00 7.21 55.6% 1 42 9,690
Qwen: Qwen3.5-27B 1.00 10.00 0.0% 0 15 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.2 9.50 10.00 100.0% 0 95 587
Qwen: Qwen3.5-27B 4.00 10.00 0.0% 0 69 0
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.2 8.00 10.00 66.7% 0 710 943
Qwen: Qwen3.5-27B 4.33 9.68 33.3% 0 706 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-5.2 1.00 1.62 66.7% 1 239 469
Qwen: Qwen3.5-27B 10.00 10.00 100.0% 0 303 0

Badilisha jozi ya ulinganisho