Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

MiniMax: MiniMax M2.5 vs OpenAI: gpt-oss-120b

Jina la modeli:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe : 2026-02-27 15:16

Muhtasari

Kipimo MiniMax: MiniMax M2.5 medium Toleo: Tarehe ya kutolewa haijulikani OpenAI: gpt-oss-120b medium Toleo: Tarehe ya kutolewa haijulikani Inapatikana bure
Nafasi #26 #25
Alama 5.64 5.64
Uthabiti 6.12 7.55
Gharama kwa matokeo 4.028 0.101
Jumla ya gharama $0.242 $0.008
Majaribio sahihi
Majaribio yenye makosa 8 7
Kiwango cha kupita kwa kila jaribio 64.3% 59.5%
Majaribio yasiyo thabiti 7 4
Tokeni za matokeo 121,297 11,407
Tokeni za hoja 203,513 26,106

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
MiniMax: MiniMax M2.5 9.33 7.89 88.9% 1 286 45,112
OpenAI: gpt-oss-120b 7.00 9.81 66.7% 0 3,463 2,077
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
MiniMax: MiniMax M2.5 5.50 5.81 83.3% 1 369 4,952
OpenAI: gpt-oss-120b 5.50 5.87 66.7% 1 241 1,114
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
MiniMax: MiniMax M2.5 1.00 4.41 22.2% 2 111,023 139,533
OpenAI: gpt-oss-120b 1.00 4.41 22.2% 2 6,018 18,520
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
MiniMax: MiniMax M2.5 7.00 6.41 66.7% 1 1,121 2,521
OpenAI: gpt-oss-120b 10.00 10.00 100.0% 0 120 1,770
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
MiniMax: MiniMax M2.5 4.33 4.79 55.6% 2 8,229 10,458
OpenAI: gpt-oss-120b 5.00 7.13 44.4% 1 1,278 1,542
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
MiniMax: MiniMax M2.5 10.00 10.00 100.0% 0 269 937
OpenAI: gpt-oss-120b 9.00 9.97 100.0% 0 287 1,083

Badilisha jozi ya ulinganisho