Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Grok 4.7 (xhigh) vs Grok Build 0.1 (medium)

Skor rata-rata hampir imbang di 6.7 vs 6.7. Grok Build 0.1 (medium) memiliki biaya benchmark lebih rendah di $1.419 vs $3.398. Grok Build 0.1 (medium) lebih cepat di 55.13s vs 123.96s, dengan tingkat keberhasilan 73.9% vs 58.0%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-03

Model yang Dibandingkan

Peringkat
#177
Total token output
572,315
Waktu respons (rata-rata)
123.96s
Total Biaya
$3.398
Peringkat
#178
Total token output
529,806
Waktu respons (rata-rata)
55.13s
Total Biaya
$1.419
Model yang direkomendasikan Grok Build 0.1 (medium)

It has the best score here (6.7), while costing about 2.4x less than Grok 4.7 (xhigh).

Perbandingan terperinci

Metrik Grok 4.7 Grok 4.7 xhigh Rilis: 2026-09-21 Grok Build 0.1 Grok Build 0.1 medium Rilis: 2026-05-21
Skor 6.7 6.7
Peringkat #177 #178
Keandalan 9.7 10.0
Konsistensi 7.8 9.6
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 73.9% 58.0%
Tes tidak stabil 6 1
Total Run 69 69
Biaya per hasil 24.272 10.910
Total Biaya $3.398 $1.419
Harga input $2.000 / 1M $1.000 / 1M
Harga output $6.000 / 1M $2.000 / 1M
Harga baca cache $0.500 / 1M $0.200 / 1M
Harga tulis cache T/A T/A
Total token input 310,049 358,616
Token output 6,094 9,651
Token penalaran 566,221 520,155
Waktu respons (rata-rata) 123.96s 55.13s
Waktu respons (maks) 775.64s 252.69s
Waktu respons (total) 2851.07s 1268.10s
Parameter ~1.7T total (~170B aktif) ~300B total (~30B aktif)
Ketersediaan Tertutup Tertutup

Harga cache berlaku untuk token input. Pembacaan memakai kembali prompt tersimpan; penulisan menyimpannya dan dapat menambah biaya. Token output menggunakan harga output.

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#177 SpaceXAI: Grok 4.7

xhigh
Biaya
$0.235
Waktu
544.2s
Token
39,247 tok

#178 SpaceXAI: Grok Build 0.1

medium
Biaya
$0.028
Waktu
81.3s
Token
14,009 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Grok 4.7 5.3 4.5 55.6% 2 414.53s 10,168 327 256,935
Grok Build 0.1 5.7 9.7 33.3% 0 108.46s 8,304 1,138 161,452

Perbandingan Cepat

Ganti Pasangan Perbandingan