Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Grok Build 0.1 (medium) vs GLM 5.3 Flash (low)

Grok Build 0.1 (medium) unggul dalam skor rata-rata dengan 6.7 vs 6.6. GLM 5.3 Flash (low) memiliki biaya benchmark lebih rendah di $0.038 vs $1.419. GLM 5.3 Flash (low) lebih cepat di 11.82s vs 55.13s, dengan tingkat keberhasilan 58.0% vs 58.0%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-03

Model yang Dibandingkan

Peringkat
#178
Total token output
529,806
Waktu respons (rata-rata)
55.13s
Total Biaya
$1.419
Peringkat
#187
Total token output
27,168
Waktu respons (rata-rata)
11.82s
Total Biaya
$0.038
Model yang direkomendasikan GLM 5.3 Flash (low)

Its score stays close to the best score here (6.6 vs 6.7), while costing about 38.2x less than Grok Build 0.1 (medium).

Perbandingan terperinci

Metrik Grok Build 0.1 Grok Build 0.1 medium Rilis: 2026-05-21 GLM 5.3 Flash GLM 5.3 Flash low Rilis: 2026-08-26
Skor 6.7 6.6
Peringkat #178 #187
Keandalan 10.0 10.0
Konsistensi 9.6 8.5
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 58.0% 58.0%
Tes tidak stabil 1 4
Total Run 69 69
Biaya per hasil 10.910 0.338
Total Biaya $1.419 $0.038
Harga input $1.000 / 1M $0.150 / 1M
Harga output $2.000 / 1M $0.500 / 1M
Harga baca cache $0.200 / 1M $0.030 / 1M
Harga tulis cache T/A T/A
Total token input 358,616 255,173
Token output 9,651 9,891
Token penalaran 520,155 17,277
Waktu respons (rata-rata) 55.13s 11.82s
Waktu respons (maks) 252.69s 90.45s
Waktu respons (total) 1268.10s 271.92s
Parameter ~300B total (~30B aktif) 320B total (18B aktif)
Ketersediaan Tertutup Sumber terbuka

Harga cache berlaku untuk token input. Pembacaan memakai kembali prompt tersimpan; penulisan menyimpannya dan dapat menambah biaya. Token output menggunakan harga output.

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#178 SpaceXAI: Grok Build 0.1

medium
Biaya
$0.028
Waktu
81.3s
Token
14,009 tok

#187 GLM 5.3 Flash

low
Biaya
$0.001
Waktu
20.7s
Token
1,482 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Grok Build 0.1 5.7 9.7 33.3% 0 108.46s 8,304 1,138 161,452
GLM 5.3 Flash 5.4 7.2 44.4% 1 10.38s 7,317 398 3,973

Perbandingan Cepat

Ganti Pasangan Perbandingan