Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Grok 4.5 (high) vs GLM 5.2 (medium)

Skor rata-rata hampir imbang di 8.2 vs 8.2. GLM 5.2 (medium) memiliki biaya benchmark lebih rendah di $0.398 vs $2.571. GLM 5.2 (medium) lebih cepat di 26.88s vs 72.50s, dengan tingkat keberhasilan 82.6% vs 81.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#82
Total token output
333,632
Waktu respons (rata-rata)
72.50s
Total Biaya
$2.571
Peringkat
#83
Total token output
81,020
Waktu respons (rata-rata)
26.88s
Total Biaya
$0.398
Model yang direkomendasikan GLM 5.2 (medium)

It has the best score here (8.2), while costing about 6.5x less than Grok 4.5 (high).

Perbandingan terperinci

Metrik Grok 4.5 Grok 4.5 high Rilis: 2026-07-08 GLM 5.2 GLM 5.2 medium Rilis: 2026-06-17
Skor 8.2 8.2
Peringkat #82 #83
Keandalan 10.0 9.6
Konsistensi 8.9 8.1
Percobaan 69/69 66/69
Tes benar
Tingkat lulus per percobaan 82.6% 81.2%
Tes tidak stabil 3 4
Total Run 69 66
Biaya per hasil 15.123 2.890
Total Biaya $2.571 $0.398
Harga input $2.000 / 1M $0.325 / 1M
Harga output $6.000 / 1M $3.990 / 1M
Total token input 342,124 227,164
Token output 7,135 20,029
Token penalaran 326,497 60,991
Waktu respons (rata-rata) 72.50s 26.88s
Waktu respons (maks) 676.83s 102.40s
Waktu respons (total) 1667.51s 591.34s
Parameter ~1.7T total (~170B aktif) 744B total (40B aktif)
Ketersediaan Tertutup Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#82 SpaceXAI: Grok 4.5

high
Biaya
$0.015
Waktu
29.7s
Token
2,737 tok

#83 GLM 5.2

medium
Biaya
$0.041
Waktu
195.8s
Token
9,287 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Grok 4.5 10.0 10.0 100.0% 0 155.16s 9,579 357 100,752
GLM 5.2 8.2 7.2 88.9% 1 40.96s 7,317 1,475 17,123

Perbandingan Cepat

Ganti Pasangan Perbandingan