Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

OpenAI: GPT-4o-mini vs StepFun: Step 3.5 Flash

Jina la modeli:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe : 2026-02-27 15:16

Muhtasari

Kipimo OpenAI: GPT-4o-mini none Toleo: Tarehe ya kutolewa haijulikani StepFun: Step 3.5 Flash medium Toleo: Tarehe ya kutolewa haijulikani Inapatikana bure
Nafasi #28 #11
Alama 4.86 7.00
Uthabiti 9.98 8.32
Gharama kwa matokeo 0.056 0.000
Jumla ya gharama $0.003 $0.000
Majaribio sahihi
Majaribio yenye makosa 9 5
Kiwango cha kupita kwa kila jaribio 35.7% 73.8%
Majaribio yasiyo thabiti 0 3
Tokeni za matokeo 949 60,502
Tokeni za hoja 0 117,044

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-4o-mini 4.00 10.00 33.3% 0 180 0
StepFun: Step 3.5 Flash 10.00 10.00 100.0% 0 13,924 17,208
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-4o-mini 10.00 10.00 100.0% 0 183 0
StepFun: Step 3.5 Flash 10.00 10.00 100.0% 0 535 11,548
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-4o-mini 1.00 10.00 0.0% 0 15 0
StepFun: Step 3.5 Flash 4.00 7.21 44.4% 1 40,942 74,237
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-4o-mini 5.50 10.00 50.0% 0 71 0
StepFun: Step 3.5 Flash 10.00 10.00 100.0% 0 2,121 3,274
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-4o-mini 4.00 9.92 0.0% 0 295 0
StepFun: Step 3.5 Flash 2.00 4.96 33.3% 2 2,705 6,975
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
OpenAI: GPT-4o-mini 10.00 10.00 100.0% 0 205 0
StepFun: Step 3.5 Flash 10.00 10.00 100.0% 0 275 3,802

Badilisha jozi ya ulinganisho