Navigasi
Advertise here

Seed-2.0-Code (medium) vs DeepSeek V4 Pro

Skor rata-rata hampir imbang di 6.8 vs 6.8. DeepSeek V4 Pro memiliki biaya benchmark lebih rendah di $0.206 vs $1.153. DeepSeek V4 Pro lebih cepat di 11.71s vs 101.05s, dengan tingkat keberhasilan 63.6% vs 47.0%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-24

Model yang Dibandingkan

Peringkat
#176
Total token output
371,443
Waktu respons (rata-rata)
101.05s
Total Biaya
$1.153
Peringkat
#180
Total token output
35,547
Waktu respons (rata-rata)
11.71s
Total Biaya
$0.206
Model yang direkomendasikan DeepSeek V4 Pro

It has the best score here (6.8), while costing about 5.6x less than Seed-2.0-Code (medium).

Perbandingan terperinci

Metrik Seed-2.0-Code Seed-2.0-Code medium Rilis: 2026-08-12 DeepSeek V4 Pro DeepSeek V4 Pro none Rilis: 2026-04-24
Skor 6.8 6.8
Peringkat #176 #180
Keandalan 7.3 10.0
Konsistensi 7.3 8.6
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 63.6% 47.0%
Tes tidak stabil 7 4
Total Run 66 66
Biaya per hasil 11.524 1.061
Total Biaya $1.153 $0.206
Harga input $0.500 / 1M $0.940 / 1M
Harga output $3.000 / 1M $1.880 / 1M
Total token input 76,036 148,078
Token output 7,267 35,547
Token penalaran 364,176 0
Waktu respons (rata-rata) 101.05s 11.71s
Waktu respons (maks) 626.58s 119.44s
Waktu respons (total) 2223.19s 257.67s
Parameter ~200B total (~20B aktif) 1.6T total (49B aktif)
Ketersediaan Tertutup Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#176 Seed-2.0-Code

medium
Biaya
$0.026
Waktu
122.8s
Token
8,630 tok

#180 DeepSeek V4 Pro

none
Reached the allocated time limit (300 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
300.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Seed-2.0-Code 8.9 7.8 88.9% 1 89.94s 8,247 496 44,062
DeepSeek V4 Pro 5.6 10.0 33.3% 0 13.38s 7,275 5,500 0

Perbandingan Cepat

Ganti Pasangan Perbandingan