Navigasi
Advertise here

Seed-2.0-Code (low) vs DeepSeek V4 Pro (high)

Seed-2.0-Code (low) unggul dalam skor rata-rata dengan 7.7 vs 7.7. DeepSeek V4 Pro (high) memiliki biaya benchmark lebih rendah di $0.447 vs $0.767. Seed-2.0-Code (low) lebih cepat di 72.24s vs 92.50s, dengan tingkat keberhasilan 77.3% vs 63.6%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-24

Model yang Dibandingkan

Peringkat
#106
Total token output
241,309
Waktu respons (rata-rata)
72.24s
Total Biaya
$0.767
Peringkat
#108
Total token output
209,444
Waktu respons (rata-rata)
92.50s
Total Biaya
$0.447
Model yang direkomendasikan DeepSeek V4 Pro (high)

Its score stays close to the best score here (7.7 vs 7.7), while costing about 1.7x less than Seed-2.0-Code (low).

Perbandingan terperinci

Metrik Seed-2.0-Code Seed-2.0-Code low Rilis: 2026-08-12 DeepSeek V4 Pro DeepSeek V4 Pro high Rilis: 2026-04-24
Skor 7.7 7.7
Peringkat #106 #108
Keandalan 8.5 10.0
Konsistensi 7.8 7.7
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 77.3% 63.6%
Tes tidak stabil 6 6
Total Run 66 66
Biaya per hasil 5.899 2.251
Total Biaya $0.767 $0.447
Harga input $0.500 / 1M $0.940 / 1M
Harga output $3.000 / 1M $1.880 / 1M
Total token input 85,657 90,757
Token output 8,142 22,928
Token penalaran 233,167 186,516
Waktu respons (rata-rata) 72.24s 92.50s
Waktu respons (maks) 485.92s 416.76s
Waktu respons (total) 1589.31s 2035.01s
Parameter ~200B total (~20B aktif) 1.6T total (49B aktif)
Ketersediaan Tertutup Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#106 Seed-2.0-Code

low
Provider returned error
Biaya
$0.000
Waktu
0.3s
Token
0 tok

#108 DeepSeek V4 Pro

high
Biaya
$0.023
Waktu
257.6s
Token
14,870 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Seed-2.0-Code 7.0 7.1 55.6% 1 109.70s 7,948 455 55,759
DeepSeek V4 Pro 6.3 8.7 33.3% 0 243.00s 5,090 383 84,580

Perbandingan Cepat

Ganti Pasangan Perbandingan