Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

OpenAI: gpt-oss-120b vs Qwen: Qwen3.5-27B

Jina la modeli:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe : 2026-02-27 15:16

Muhtasari

Kipimo OpenAI: gpt-oss-120b medium Toleo: Tarehe ya kutolewa haijulikani Inapatikana bure Qwen: Qwen3.5-27B medium Toleo: Tarehe ya kutolewa haijulikani
Nafasi #25 #5
Alama 5.64 8.55
Uthabiti 7.55 9.55
Gharama kwa matokeo 0.101 2.950
Jumla ya gharama $0.008 $0.325
Majaribio sahihi
Majaribio yenye makosa 7 3
Kiwango cha kupita kwa kila jaribio 59.5% 83.3%
Majaribio yasiyo thabiti 4 1
Tokeni za matokeo 11,407 1,091
Tokeni za hoja 26,106 131,807

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: gpt-oss-120b 7.00 9.81 66.7% 0 3,463 2,077
Qwen: Qwen3.5-27B 10.00 10.00 100.0% 0 102 8,956
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: gpt-oss-120b 5.50 5.87 66.7% 1 241 1,114
Qwen: Qwen3.5-27B 9.88 10.00 100.0% 0 270 16,150
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: gpt-oss-120b 1.00 4.41 22.2% 2 6,018 18,520
Qwen: Qwen3.5-27B 4.00 10.00 33.3% 0 43 52,368
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: gpt-oss-120b 10.00 10.00 100.0% 0 120 1,770
Qwen: Qwen3.5-27B 9.00 6.88 83.3% 1 97 11,638
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: gpt-oss-120b 5.00 7.13 44.4% 1 1,278 1,542
Qwen: Qwen3.5-27B 10.00 10.00 100.0% 0 231 41,372
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: gpt-oss-120b 9.00 9.97 100.0% 0 287 1,083
Qwen: Qwen3.5-27B 10.00 10.00 100.0% 0 348 1,323

Badilisha jozi ya ulinganisho