Navigasi
AI BENCHY
Advertise here

DeepSeek V4 Flash 0731 (medium) vs GLM 5.3 Flash (high)

Skor rata-rata hampir imbang di 7.3 vs 7.3. GLM 5.3 Flash (high) memiliki biaya benchmark lebih rendah di $0.029 vs $0.046. GLM 5.3 Flash (high) lebih cepat di 25.36s vs 84.66s, dengan tingkat keberhasilan 71.2% vs 66.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-08-26

Peringkat
#105
Total token output
550,855
Waktu respons (rata-rata)
84.66s
Total Biaya
$0.046
Peringkat
#102
Total token output
78,318
Waktu respons (rata-rata)
25.36s
Total Biaya
$0.029
Model yang direkomendasikan GLM 5.3 Flash (high)

It has the best score here (7.3), while costing about 1.6x less than DeepSeek V4 Flash 0731 (medium).

Perbandingan terperinci

Metrik DeepSeek V4 Flash 0731 DeepSeek V4 Flash 0731 medium Rilis: 2026-08-01 GLM 5.3 Flash GLM 5.3 Flash high Rilis: 2026-08-26
Skor 7.3 7.3
Peringkat #105 #102
Keandalan 10.0 10.0
Konsistensi 7.1 8.9
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 71.2% 66.7%
Tes tidak stabil 8 3
Total Run 66 66
Biaya per hasil 0.960 0.216
Total Biaya $0.046 $0.029
Harga input $0.060 / 1M $0.075 / 1M
Harga output $0.120 / 1M $0.250 / 1M
Total token input 89,819 113,003
Token output 58,398 11,558
Token penalaran 492,457 66,760
Waktu respons (rata-rata) 84.66s 25.36s
Waktu respons (maks) 288.78s 163.41s
Waktu respons (total) 1862.46s 558.01s
Parameter 284B total (13B aktif) 320B total (18B aktif)
Ketersediaan Tertutup Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#105 DeepSeek V4 Flash 0731

medium
Biaya
$0.004
Waktu
93.3s
Token
13,252 tok

#102 GLM 5.3 Flash

high
Biaya
$0.003
Waktu
186.7s
Token
10,486 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V4 Flash 0731 6.4 4.4 77.8% 2 198.78s 6,032 4,109 160,154
GLM 5.3 Flash 8.2 7.2 88.9% 1 19.62s 7,317 373 9,973

Perbandingan Cepat

Ganti Pasangan Perbandingan