Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

Qwen: Qwen3.5-27B vs StepFun: Step 3.7 Flash

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-29

Kipimo Qwen3.5-27B Qwen3.5-27B medium Toleo: 2026-02-24 Step 3.7 Flash Step 3.7 Flash medium Toleo: 2026-05-29
Alama 7.9 7.9
Nafasi #28 #32
Uaminifu 10.0 9.9
Uthabiti 8.9 9.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 73.3% 71.7%
Majaribio yasiyo thabiti 3 2
Jumla ya uendeshaji 60 58
Gharama kwa matokeo 4.532 2.663
Jumla ya gharama $0.590 $0.347
Bei ya ingizo $0.195 / 1M $0.200 / 1M
Bei ya toleo $1.560 / 1M $1.150 / 1M
Tokeni za matokeo 2,569 294,481
Tokeni za hoja 304,894 0
Muda wa majibu (wastani) 60.09s 18.32s
Muda wa majibu (upeo) 177.36s 113.98s
Muda wa majibu (jumla) 1201.89s 366.45s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 8.7 7.9 91.7% 1 19.75s 569 31,505
Step 3.7 Flash 8.7 7.9 91.7% 1 9.65s 32,185 0
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 7.0 9.8 50.0% 0 123.86s 416 64,993
Step 3.7 Flash 8.2 6.7 83.3% 1 10.64s 19,320 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 10.0 10.0 100.0% 0 163.96s 483 9,991
Step 3.7 Flash 10.0 10.0 100.0% 0 9.06s 7,106 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 10.0 10.0 100.0% 0 30.26s 270 16,150
Step 3.7 Flash 10.0 10.0 100.0% 0 2.75s 3,020 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 5.3 10.0 33.3% 0 79.53s 43 52,368
Step 3.7 Flash 7.7 10.0 66.7% 0 48.27s 70,347 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 6.1 3.1 66.7% 1 101.41s 70 23,147
Step 3.7 Flash 4.0 10.0 0.0% 0 6.85s 3,987 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 10.0 10.0 100.0% 0 19.66s 97 11,638
Step 3.7 Flash 9.8 10.0 100.0% 0 1.83s 2,166 0
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 8.2 7.7 77.8% 1 59.60s 242 70,096
Step 3.7 Flash 5.7 9.9 33.3% 0 6.19s 15,071 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 10.0 10.0 100.0% 0 7.45s 348 1,323
Step 3.7 Flash 10.0 10.0 100.0% 0 4.16s 2,115 0
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Qwen3.5-27B 3.0 10.0 0.0% 0 85.11s 31 23,683
Step 3.7 Flash 3.0 10.0 0.0% 0 113.98s 139,164 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho