Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

DeepSeek V3.2 (medium) vs Qwen3.7 Flash (medium)

DeepSeek V3.2 (medium) unggul dalam skor rata-rata dengan 7.0 vs 7.0. Qwen3.7 Flash (medium) memiliki biaya benchmark lebih rendah di $0.064 vs $0.078. Qwen3.7 Flash (medium) lebih cepat di 45.90s vs 68.62s, dengan tingkat keberhasilan 65.2% vs 65.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-07-28

Peringkat
#86
Total token output
128,848
Waktu respons (rata-rata)
68.62s
Total Biaya
$0.078
Peringkat
#91
Total token output
462,623
Waktu respons (rata-rata)
45.90s
Total Biaya
$0.064
Model yang direkomendasikan Qwen3.7 Flash (medium)

It offers the best overall trade-off: a competitive score (7.0), lower cost than DeepSeek V3.2 (medium), and balanced response time.

Perbandingan terperinci

Metrik DeepSeek V3.2 DeepSeek V3.2 medium Rilis: 2025-12-01 Qwen3.7 Flash Qwen3.7 Flash medium Rilis: 2026-07-28
Skor 7.0 7.0
Peringkat #86 #91
Keandalan 10.0 10.0
Konsistensi 7.4 7.5
Tes benar
Tingkat lulus per percobaan 65.2% 65.2%
Tes tidak stabil 7 7
Total Run 66 66
Biaya per hasil 0.671 0.636
Total Biaya $0.078 $0.064
Harga input $0.269 / 1M $0.030 / 1M
Harga output $0.400 / 1M $0.130 / 1M
Total token input 101,047 114,468
Token output 11,834 12,786
Token penalaran 117,014 449,837
Waktu respons (rata-rata) 68.62s 45.90s
Waktu respons (maks) 376.10s 593.64s
Waktu respons (total) 1509.53s 1009.84s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#86 DeepSeek V3.2

medium
Biaya
$0.001
Waktu
53.6s
Token
1,932 tok

#91 Qwen3.7 Flash

medium
SVG tidak valid
Biaya
$0.000
Waktu
600.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V3.2 6.0 7.2 55.6% 1 248.68s 5,717 649 52,014
Qwen3.7 Flash 6.2 6.9 55.6% 1 46.78s 7,893 563 61,902

Perbandingan Cepat

Ganti Pasangan Perbandingan