Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

StepFun: Step 3.7 Flash vs Z.ai: GLM 5.1

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-29

Kipimo Step 3.7 Flash Step 3.7 Flash low Toleo: 2026-05-29 GLM 5.1 GLM 5.1 medium Toleo: 2026-04-07
Alama 7.4 7.4
Nafasi #60 #56
Uaminifu 10.0 5.0
Uthabiti 8.7 8.3
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 68.3% 71.7%
Majaribio yasiyo thabiti 3 4
Jumla ya uendeshaji 60 60
Gharama kwa matokeo 2.796 2.382
Jumla ya gharama $0.336 $0.286
Bei ya ingizo $0.200 / 1M $0.980 / 1M
Bei ya toleo $1.150 / 1M $3.080 / 1M
Tokeni za matokeo 285,209 11,511
Tokeni za hoja 0 71,979
Muda wa majibu (wastani) 16.06s 33.45s
Muda wa majibu (upeo) 124.75s 172.60s
Muda wa majibu (jumla) 321.11s 635.63s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 8.7 7.9 91.7% 1 4.02s 10,896 0
GLM 5.1 10.0 10.0 100.0% 0 8.31s 401 5,122
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 9.43s 14,569 0
GLM 5.1 4.7 1.6 66.7% 2 145.56s 4,727 34,384
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 7.98s 6,426 0
GLM 5.1 9.5 10.0 100.0% 0 43.11s 327 4,206
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 7.3 5.8 83.3% 1 2.29s 2,667 0
GLM 5.1 10.0 10.0 100.0% 0 9.33s 991 4,552
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 5.3 7.2 44.4% 1 43.31s 104,487 0
GLM 5.1 5.3 10.0 33.3% 0 29.77s 969 11,314
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 3.4 9.3 0.0% 0 7.00s 4,604 0
GLM 5.1 10.0 10.0 100.0% 0 20.95s 2,875 2,875
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 9.8 10.0 100.0% 0 1.58s 1,857 0
GLM 5.1 6.4 5.8 66.7% 1 7.47s 204 1,617
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 5.5 9.9 33.3% 0 1.84s 3,564 0
GLM 5.1 8.2 7.2 88.9% 1 31.64s 935 5,730
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 3.25s 1,360 0
GLM 5.1 3.0 10.0 0.0% 0 0ms 0 0
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 3.0 10.0 0.0% 0 124.75s 134,779 0
GLM 5.1 3.0 10.0 0.0% 0 29.40s 82 2,179

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho