Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

StepFun: Step 3.7 Flash vs xAI: Grok 4.3

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-29

Kipimo Step 3.7 Flash Step 3.7 Flash low Toleo: 2026-05-29 Grok 4.3 Grok 4.3 medium Toleo: 2026-05-01
Alama 7.4 7.8
Nafasi #60 #36
Uaminifu 10.0 10.0
Uthabiti 8.7 8.4
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 68.3% 75.0%
Majaribio yasiyo thabiti 3 4
Jumla ya uendeshaji 60 60
Gharama kwa matokeo 2.796 4.557
Jumla ya gharama $0.336 $0.593
Bei ya ingizo $0.200 / 1M $1.250 / 1M
Bei ya toleo $1.150 / 1M $2.500 / 1M
Tokeni za matokeo 285,209 1,485
Tokeni za hoja 0 214,710
Muda wa majibu (wastani) 16.06s 49.23s
Muda wa majibu (upeo) 124.75s 216.69s
Muda wa majibu (jumla) 321.11s 984.52s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 8.7 7.9 91.7% 1 4.02s 10,896 0
Grok 4.3 10.0 10.0 100.0% 0 8.83s 88 8,207
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 9.43s 14,569 0
Grok 4.3 7.4 6.5 66.7% 1 55.26s 532 24,554
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 7.98s 6,426 0
Grok 4.3 10.0 10.0 100.0% 0 63.99s 234 15,301
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 7.3 5.8 83.3% 1 2.29s 2,667 0
Grok 4.3 10.0 10.0 100.0% 0 18.97s 180 9,546
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 5.3 7.2 44.4% 1 43.31s 104,487 0
Grok 4.3 5.3 7.2 44.4% 1 181.74s 14 111,300
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 3.4 9.3 0.0% 0 7.00s 4,604 0
Grok 4.3 5.4 2.5 66.7% 1 24.70s 70 5,020
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 9.8 10.0 100.0% 0 1.58s 1,857 0
Grok 4.3 9.8 10.0 100.0% 0 18.58s 57 8,713
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 5.5 9.9 33.3% 0 1.84s 3,564 0
Grok 4.3 5.9 7.2 55.6% 1 22.52s 128 14,468
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 3.25s 1,360 0
Grok 4.3 10.0 10.0 100.0% 0 17.66s 168 4,615
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 3.0 10.0 0.0% 0 124.75s 134,779 0
Grok 4.3 3.0 10.0 0.0% 0 44.47s 14 12,986

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho