Navigasi
Advertise here

Claude Opus 4.6 (medium) vs Seed-2.0-Code (low)

Skor rata-rata hampir imbang di 7.7 vs 7.7. Seed-2.0-Code (low) memiliki biaya benchmark lebih rendah di $0.767 vs $2.961. Claude Opus 4.6 (medium) lebih cepat di 32.23s vs 72.24s, dengan tingkat keberhasilan 63.6% vs 77.3%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-24

Model yang Dibandingkan

Peringkat
#107
Total token output
96,722
Waktu respons (rata-rata)
32.23s
Total Biaya
$2.961
Peringkat
#106
Total token output
241,309
Waktu respons (rata-rata)
72.24s
Total Biaya
$0.767
Model yang direkomendasikan Seed-2.0-Code (low)

It has the best score here (7.7), while costing about 3.9x less than Claude Opus 4.6 (medium).

Perbandingan terperinci

Metrik Claude Opus 4.6 Claude Opus 4.6 medium Rilis: 2026-02-05 Seed-2.0-Code Seed-2.0-Code low Rilis: 2026-08-12
Skor 7.7 7.7
Peringkat #107 #106
Keandalan 10.0 8.5
Konsistensi 8.8 7.8
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 63.6% 77.3%
Tes tidak stabil 3 6
Total Run 66 66
Biaya per hasil 22.777 5.899
Total Biaya $2.961 $0.767
Harga input $5.000 / 1M $0.500 / 1M
Harga output $25.000 / 1M $3.000 / 1M
Total token input 108,573 85,657
Token output 69,994 8,142
Token penalaran 26,728 233,167
Waktu respons (rata-rata) 32.23s 72.24s
Waktu respons (maks) 151.51s 485.92s
Waktu respons (total) 515.65s 1589.31s
Parameter ~5T total (~500B aktif) ~200B total (~20B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#107 Claude Opus 4.6

medium
Reached the allocated time limit (300 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
300.0s
Token
0 tok

#106 Seed-2.0-Code

low
Provider returned error
Biaya
$0.000
Waktu
0.3s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Claude Opus 4.6 5.7 7.1 44.4% 1 30.10s 8,522 13,057 4,121
Seed-2.0-Code 7.0 7.1 55.6% 1 109.70s 7,948 455 55,759

Perbandingan Cepat

Ganti Pasangan Perbandingan