Navigasi
Advertise here

Seed-2.0-Lite vs Step 5 Preview (medium)

Step 5 Preview (medium) unggul dalam skor rata-rata dengan 6.1 vs 6.0. Seed-2.0-Lite memiliki biaya benchmark lebih rendah di $0.114 vs $4.136. Seed-2.0-Lite lebih cepat di 7.49s vs 136.10s, dengan tingkat keberhasilan 39.1% vs 53.6%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-09

Model yang Dibandingkan

Peringkat
#258
Total token output
17,558
Waktu respons (rata-rata)
7.49s
Total Biaya
$0.114
Peringkat
#250
Total token output
1,434,621
Waktu respons (rata-rata)
136.10s
Total Biaya
$4.136
Model yang direkomendasikan Seed-2.0-Lite

Its score stays close to the best score here (6.0 vs 6.1), while costing about 36.4x less than Step 5 Preview (medium).

Perbandingan terperinci

Metrik Seed-2.0-Lite Seed-2.0-Lite none Rilis: 2026-02-14 Step 5 Preview Step 5 Preview medium Rilis: 2026-10-09
Skor 6.0 6.1
Peringkat #258 #250
Keandalan 10.0 9.6
Konsistensi 8.9 8.2
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 39.1% 53.6%
Tes tidak stabil 3 5
Total Run 69 69
Biaya per hasil 1.421 41.354
Total Biaya $0.114 $4.136
Harga input $0.250 / 1M $1.000 / 1M
Harga output $2.000 / 1M $2.700 / 1M
Harga baca cache T/A $0.050 / 1M
Harga tulis cache T/A T/A
Total token input 314,113 261,842
Token output 17,558 1,434,621
Token penalaran 0 0
Waktu respons (rata-rata) 7.49s 136.10s
Waktu respons (maks) 75.99s 438.09s
Waktu respons (total) 172.19s 3130.27s
Parameter ~200B total (~20B aktif) 600B total (27B aktif)
Ketersediaan Tertutup Tertutup

Harga cache berlaku untuk token input. Pembacaan memakai kembali prompt tersimpan; penulisan menyimpannya dan dapat menambah biaya. Token output menggunakan harga output.

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#258 Seed-2.0-Lite

none
Biaya
$0.005
Waktu
83.8s
Token
2,311 tok

#250 Step 5 Preview

medium
Biaya
$0.042
Waktu
90.3s
Token
15,574 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Seed-2.0-Lite 5.6 10.0 33.3% 0 2.83s 8,215 410 0
Step 5 Preview 2.8 7.2 11.1% 1 363.90s 7,428 546,725 0

Ganti Pasangan Perbandingan