Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Seed-2.0-Code (medium) vs GLM 5.2

Seed-2.0-Code (medium) unggul dalam skor rata-rata dengan 6.8 vs 6.7. GLM 5.2 memiliki biaya benchmark lebih rendah di $0.078 vs $1.153. GLM 5.2 lebih cepat di 9.65s vs 101.05s, dengan tingkat keberhasilan 63.6% vs 63.6%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-08-15

Peringkat
#129
Total token output
371,443
Waktu respons (rata-rata)
101.05s
Total Biaya
$1.153
Peringkat
#136
Total token output
14,344
Waktu respons (rata-rata)
9.65s
Total Biaya
$0.078
Model yang direkomendasikan GLM 5.2

Its score stays close to the best score here (6.7 vs 6.8), while costing about 14.9x less than Seed-2.0-Code (medium).

Perbandingan terperinci

Metrik Seed-2.0-Code Seed-2.0-Code medium Rilis: 2026-08-12 GLM 5.2 GLM 5.2 none Rilis: 2026-06-17
Skor 6.8 6.7
Peringkat #129 #136
Keandalan 7.3 10.0
Konsistensi 7.3 9.2
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 63.6% 63.6%
Tes tidak stabil 7 2
Total Run 66 66
Biaya per hasil 11.524 1.312
Total Biaya $1.153 $0.078
Harga input $0.500 / 1M $0.490 / 1M
Harga output $3.000 / 1M $1.540 / 1M
Total token input 76,036 112,458
Token output 7,267 14,344
Token penalaran 364,176 0
Waktu respons (rata-rata) 101.05s 9.65s
Waktu respons (maks) 626.58s 79.65s
Waktu respons (total) 2223.19s 212.38s
Parameter ~200B total (~20B aktif) 744B total (40B aktif)
Ketersediaan Tertutup Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#129 Seed-2.0-Code

medium
Biaya
$0.026
Waktu
122.8s
Token
8,630 tok

#136 GLM 5.2

none
SVG tidak valid
Biaya
$0.033
Waktu
87.7s
Token
7,455 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Seed-2.0-Code 8.9 7.8 88.9% 1 89.94s 8,247 496 44,062
GLM 5.2 3.7 9.5 0.0% 0 7.55s 7,263 1,958 0

Perbandingan Cepat

Ganti Pasangan Perbandingan