Navigasi
Advertise here

DeepSeek V4.1 Flash (low) vs Qwen3.8 Max (low)

Qwen3.8 Max (low) unggul dalam skor rata-rata dengan 8.6 vs 8.5. DeepSeek V4.1 Flash (low) memiliki biaya benchmark lebih rendah di $0.224 vs $0.702. Qwen3.8 Max (low) lebih cepat di 21.83s vs 26.77s, dengan tingkat keberhasilan 74.2% vs 80.3%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-10

Model yang Dibandingkan

Peringkat
#51
Total token output
347,963
Waktu respons (rata-rata)
26.77s
Total Biaya
$0.224
Peringkat
#49
Total token output
83,821
Waktu respons (rata-rata)
21.83s
Total Biaya
$0.702
Model yang direkomendasikan DeepSeek V4.1 Flash (low)

Its score stays close to the best score here (8.5 vs 8.6), while costing about 3.1x less than Qwen3.8 Max (low).

Perbandingan terperinci

Metrik DeepSeek V4.1 Flash DeepSeek V4.1 Flash low Rilis: 2026-09-10 Qwen3.8 Max Qwen3.8 Max low Rilis: 2026-08-04
Skor 8.5 8.6
Peringkat #51 #49
Keandalan 9.5 10.0
Konsistensi 8.9 9.0
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 74.2% 80.3%
Tes tidak stabil 3 3
Total Run 66 66
Biaya per hasil 1.492 4.384
Total Biaya $0.224 $0.702
Harga input $0.150 / 1M $2.000 / 1M
Harga output $0.600 / 1M $6.000 / 1M
Total token input 99,715 99,181
Token output 6,381 6,132
Token penalaran 341,582 77,689
Waktu respons (rata-rata) 26.77s 21.83s
Waktu respons (maks) 149.84s 155.75s
Waktu respons (total) 589.01s 480.37s
Parameter 748B total (16B aktif) 2.4T total (~100B aktif)
Ketersediaan Sumber terbuka Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#51 DeepSeek V4.1 Flash

low
Biaya
$0.017
Waktu
43.6s
Token
14,069 tok

#49 Qwen3.8 Max

low
Biaya
$0.026
Waktu
68.5s
Token
4,280 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V4.1 Flash 10.0 10.0 100.0% 0 40.66s 7,509 378 81,978
Qwen3.8 Max 8.4 7.4 88.9% 1 28.69s 8,127 512 15,952

Perbandingan Cepat

Ganti Pasangan Perbandingan