Navigasi
Advertise here

DeepSeek V4.1 Flash vs Qwen3.8 Flash Next

Skor rata-rata hampir imbang di 5.8 vs 5.8. Qwen3.8 Flash Next memiliki biaya benchmark lebih rendah di ~$0.011 vs $0.453. Qwen3.8 Flash Next lebih cepat di 5.71s vs 16.63s, dengan tingkat keberhasilan 36.2% vs 43.5%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-05

Model yang Dibandingkan

Peringkat
#260
Total token output
292,779
Waktu respons (rata-rata)
16.63s
Total Biaya
$0.453
Peringkat
#258
Total token output
9,077
Waktu respons (rata-rata)
5.71s
Total Biaya
~$0.011
Model yang direkomendasikan Qwen3.8 Flash Next

It has the best score here (5.8), while costing about 41.8x less than DeepSeek V4.1 Flash.

Perbandingan terperinci

Metrik DeepSeek V4.1 Flash DeepSeek V4.1 Flash none Rilis: 2026-09-10 Qwen3.8 Flash Next Qwen3.8 Flash Next none Rilis: 2026-10-04
Skor 5.8 5.8
Peringkat #260 #258
Keandalan 9.7 10.0
Konsistensi 9.1 9.0
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 36.2% 43.5%
Tes tidak stabil 3 3
Total Run 69 69
Biaya per hasil 2.806 ~0.136
Total Biaya $0.453 ~$0.011
Harga input $0.003 / 1M T/A
Harga output $2.400 / 1M T/A
Harga baca cache $0.003 / 1M T/A
Harga tulis cache T/A T/A
Total token input 337,756 266,004
Token output 292,779 9,077
Token penalaran 0 0
Waktu respons (rata-rata) 16.63s 5.71s
Waktu respons (maks) 247.66s 65.25s
Waktu respons (total) 382.53s 131.36s
Parameter 748B total (16B aktif) 180B total (6B aktif)
Ketersediaan Sumber terbuka Bobot tersedia

Harga cache berlaku untuk token input. Pembacaan memakai kembali prompt tersimpan; penulisan menyimpannya dan dapat menambah biaya. Token output menggunakan harga output.

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#260 DeepSeek V4.1 Flash

none
Biaya
$0.021
Waktu
55.4s
Token
17,181 tok

#258 Qwen3.8 Flash Next

none
Biaya
~$0.001
Waktu
27.5s
Token
2,983 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V4.1 Flash 5.5 10.0 33.3% 0 83.66s 7,284 263,438 0
Qwen3.8 Flash Next 5.5 10.0 33.3% 0 1.64s 7,911 669 0

Perbandingan Cepat

Ganti Pasangan Perbandingan

DeepSeek V4.1 FlashnonevsLing 3.0 FlashhighDeepSeek V4.1 FlashnonevsRing 2.6 1tmediumTrinity Large ThinkinghighvsDeepSeek V4.1 FlashnoneDeepSeek V4.1 FlashnonevsGPT-5 NanomediumLing 3.0 FlashhighvsQwen3.8 Flash Nextnonegpt-oss-120bmediumvsQwen3.8 Flash NextnoneDeepSeek V4.1 FlashnonevsDots 3 Note PreviewmediumTersedia gratisRing 2.6 1tmediumvsQwen3.8 Flash NextnoneTrinity Large ThinkinghighvsQwen3.8 Flash NextnoneGemini 3.1 Flash LiteminimalvsQwen3.8 Flash NextnoneGPT-5 NanomediumvsQwen3.8 Flash NextnoneDeepSeek V4.1 FlashnonevsSpace Bunny Alphalow