Navigasi
Advertise here

DeepSeek V4 Flash 0731 (low) vs GPT-5 Mini (medium)

GPT-5 Mini (medium) unggul dalam skor rata-rata dengan 8.1 vs 8.0. DeepSeek V4 Flash 0731 (low) memiliki biaya benchmark lebih rendah di $0.044 vs $0.241. GPT-5 Mini (medium) lebih cepat di 27.64s vs 77.55s, dengan tingkat keberhasilan 71.2% vs 65.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-10

Model yang Dibandingkan

Peringkat
#73
Total token output
385,989
Waktu respons (rata-rata)
77.55s
Total Biaya
$0.044
Peringkat
#69
Total token output
108,019
Waktu respons (rata-rata)
27.64s
Total Biaya
$0.241
Model yang direkomendasikan DeepSeek V4 Flash 0731 (low)

Its score stays close to the best score here (8.0 vs 8.1), while costing about 5.5x less than GPT-5 Mini (medium).

Perbandingan terperinci

Metrik DeepSeek V4 Flash 0731 DeepSeek V4 Flash 0731 low Rilis: 2026-08-01 GPT-5 Mini GPT-5 Mini medium Rilis: 2025-08-07
Skor 8.0 8.1
Peringkat #73 #69
Keandalan 10.0 10.0
Konsistensi 8.3 8.4
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 71.2% 65.2%
Tes tidak stabil 5 4
Total Run 66 66
Biaya per hasil 0.553 2.006
Total Biaya $0.044 $0.241
Harga input $0.065 / 1M $0.250 / 1M
Harga output $0.180 / 1M $2.000 / 1M
Total token input 93,345 98,383
Token output 35,723 14,409
Token penalaran 350,266 93,610
Waktu respons (rata-rata) 77.55s 27.64s
Waktu respons (maks) 438.65s 111.48s
Waktu respons (total) 1706.18s 608.02s
Parameter 284B total (13B aktif) ~400B total (~17B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#73 DeepSeek V4 Flash 0731

low
No output was saved. The original provider response or failure reason is unavailable.
Biaya
$0.000
Waktu
78.1s
Token
8,090 tok

#69 GPT-5 Mini

medium
Biaya
$0.007
Waktu
42.9s
Token
3,432 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V4 Flash 0731 7.3 5.1 77.8% 2 163.31s 5,953 782 86,028
GPT-5 Mini 10.0 10.0 100.0% 0 27.63s 7,302 658 17,152

Perbandingan Cepat

Ganti Pasangan Perbandingan