Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Grok 4.7 (high) vs GLM 5.3 Flash (low)

Skor rata-rata hampir imbang di 6.7 vs 6.6. GLM 5.3 Flash (low) memiliki biaya benchmark lebih rendah di $0.052 vs $3.582. GLM 5.3 Flash (low) lebih cepat di 11.82s vs 122.19s, dengan tingkat keberhasilan 72.5% vs 58.0%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#179
Total token output
487,184
Waktu respons (rata-rata)
122.19s
Total Biaya
$3.582
Peringkat
#184
Total token output
27,168
Waktu respons (rata-rata)
11.82s
Total Biaya
$0.052
Model yang direkomendasikan GLM 5.3 Flash (low)

It has the best score here (6.6), while costing about 69.1x less than Grok 4.7 (high).

Perbandingan terperinci

Metrik Grok 4.7 Grok 4.7 high Rilis: 2026-09-21 GLM 5.3 Flash GLM 5.3 Flash low Rilis: 2026-08-26
Skor 6.7 6.6
Peringkat #179 #184
Keandalan 9.6 10.0
Konsistensi 7.7 8.5
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 72.5% 58.0%
Tes tidak stabil 7 4
Total Run 69 69
Biaya per hasil 23.360 0.338
Total Biaya $3.582 $0.052
Harga input $2.000 / 1M $0.150 / 1M
Harga output $6.000 / 1M $0.500 / 1M
Total token input 329,276 255,173
Token output 8,172 9,891
Token penalaran 479,012 17,277
Waktu respons (rata-rata) 122.19s 11.82s
Waktu respons (maks) 573.57s 90.45s
Waktu respons (total) 2810.40s 271.92s
Parameter ~1.7T total (~170B aktif) 320B total (18B aktif)
Ketersediaan Tertutup Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#179 SpaceXAI: Grok 4.7

high
Biaya
$0.110
Waktu
301.3s
Token
23,143 tok

#184 GLM 5.3 Flash

low
Biaya
$0.001
Waktu
20.7s
Token
1,482 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Grok 4.7 6.0 4.6 66.7% 2 335.92s 10,266 353 199,976
GLM 5.3 Flash 5.4 7.2 44.4% 1 10.38s 7,317 398 3,973

Perbandingan Cepat

Ganti Pasangan Perbandingan