Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

DeepSeek V4 Pro (high) vs Qwen3.8 27B (medium)

Qwen3.8 27B (medium) unggul dalam skor rata-rata dengan 7.8 vs 7.7. Biaya perangkat keras lokal tidak diukur, sehingga perbandingan biaya tidak tersedia. Qwen3.8 27B (medium) lebih cepat di 33.05s vs 92.50s, dengan tingkat keberhasilan 63.6% vs 66.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-08-15

Peringkat
#75
Total token output
209,444
Waktu respons (rata-rata)
92.50s
Total Biaya
$0.555
Peringkat
#68
Total token output
136,162
Waktu respons (rata-rata)
33.05s
Total Biaya
T/A
Model yang direkomendasikan DeepSeek V4 Pro (high)

It has the best overall balance of score, reliability, cost, and response time in this comparison.

Perbandingan terperinci

Metrik DeepSeek V4 Pro DeepSeek V4 Pro high Rilis: 2026-04-24 Qwen3.8 27B Qwen3.8 27B medium Rilis: 2026-08-14
Skor 7.7 7.8
Peringkat #75 #68
Keandalan 10.0 9.6
Konsistensi 7.7 9.7
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 63.6% 66.7%
Tes tidak stabil 6 1
Total Run 66 66
Biaya per hasil 2.251 T/A
Total Biaya $0.555 T/A
Harga input $1.169 / 1M T/A
Harga output $2.337 / 1M T/A
Total token input 90,757 98,266
Token output 22,928 1,861
Token penalaran 186,516 134,301
Waktu respons (rata-rata) 92.50s 33.05s
Waktu respons (maks) 416.76s 214.63s
Waktu respons (total) 2035.01s 694.04s
Parameter 1.6T total (49B aktif) 27.3B
Ketersediaan Sumber terbuka Bobot tersedia

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#75 DeepSeek V4 Pro

high
Biaya
$0.023
Waktu
257.6s
Token
14,870 tok

#68 Qwen3.8 27B

medium
Biaya
T/A
Waktu
43.4s
Token
2,956 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V4 Pro 6.3 8.7 33.3% 0 243.00s 5,090 383 84,580
Qwen3.8 27B 10.0 10.0 100.0% 0 40.50s 7,893 585 26,808

Perbandingan Cepat

Ganti Pasangan Perbandingan