Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

StepFun: Step 3.7 Flash vs xAI: Grok Build 0.1

Muhtasari

Ulinganisho wa benchmark Step 3.7 Flash vs Grok Build 0.1: Grok Build 0.1 inaongoza kwa average score: 7.6 vs 7.1. Grok Build 0.1 ina gharama ya chini ya benchmark: $0.927 vs $1.148. Grok Build 0.1 ni ya haraka zaidi: 49.90s vs 64.46s, na pass rates 63.5% vs 61.9%.

Muundo unaopendekezwa: Grok Build 0.1 - It has the strongest score in this comparison (7.6) and the best overall balance of cost and response time across all 2 models.

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-07-02

Kipimo Step 3.7 Flash Step 3.7 Flash high Toleo: 2026-05-29 Grok Build 0.1 Grok Build 0.1 medium Toleo: 2026-05-21
Alama 7.1 7.6
Nafasi #65 #44
Uaminifu 10.0 10.0
Uthabiti 8.2 9.9
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 63.5% 61.9%
Majaribio yasiyo thabiti 4 0
Jumla ya uendeshaji 63 63
Gharama kwa matokeo 10.434 7.124
Jumla ya gharama $1.148 $0.927
Bei ya ingizo $0.200 / 1M $1.000 / 1M
Bei ya toleo $1.150 / 1M $2.000 / 1M
Jumla ya tokeni za ingizo 38,391 44,418
Tokeni za matokeo 991,355 2,782
Tokeni za hoja 0 438,018
Muda wa majibu (wastani) 64.46s 49.90s
Muda wa majibu (upeo) 364.99s 252.69s
Muda wa majibu (jumla) 1353.57s 1047.92s

Onyesho la kizazi

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#65 Step 3.7 Flash

high
Gharama
$0.007
Muda
63.6s
Tokeni
6,030 tok

#44 xAI: Grok Build 0.1

medium
Gharama
$0.028
Muda
81.3s
Tokeni
14,009 tok

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 13.40s 696 42,656 0
Grok Build 0.1 8.3 10.0 75.0% 0 7.43s 2,010 220 12,162
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 4.0 6.0 22.2% 1 206.21s 6,057 327,340 0
Grok Build 0.1 5.7 9.7 33.3% 0 108.46s 8,304 1,138 161,452
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 13.01s 13,638 8,802 0
Grok Build 0.1 10.0 10.0 100.0% 0 32.81s 12,909 231 16,917
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 14.72s 7,368 23,113 0
Grok Build 0.1 10.0 10.0 100.0% 0 10.72s 7,761 180 8,876
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 4.1 4.4 44.5% 2 149.64s 783 410,502 0
Grok Build 0.1 5.3 10.0 33.3% 0 158.00s 1,764 492 175,294
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 5.5 10.0 0.0% 0 4.17s 510 2,862 0
Grok Build 0.1 4.4 9.9 0.0% 0 18.41s 825 76 6,345
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 9.8 10.0 100.0% 0 1.52s 705 2,010 0
Grok Build 0.1 9.8 10.0 100.0% 0 12.36s 1,362 57 9,599
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 5.3 7.2 44.4% 1 10.22s 711 25,422 0
Grok Build 0.1 7.7 10.0 66.7% 0 18.26s 1,689 195 20,841
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 10.0 10.0 100.0% 0 2.79s 7,701 1,172 0
Grok Build 0.1 10.0 10.0 100.0% 0 13.12s 7,263 180 4,969
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Step 3.7 Flash 3.0 10.0 0.0% 0 149.34s 222 147,476 0
Grok Build 0.1 3.0 10.0 0.0% 0 53.51s 531 13 21,563

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho