Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Seed-2.0-Lite (medium) vs Qwen3.8 27B (low)

Qwen3.8 27B (low) unggul dalam skor rata-rata dengan 7.9 vs 7.8. Biaya perangkat keras lokal tidak diukur, sehingga perbandingan biaya tidak tersedia. Qwen3.8 27B (low) lebih cepat di 39.11s vs 47.94s, dengan tingkat keberhasilan 69.7% vs 68.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-08-15

Peringkat
#69
Total token output
101,296
Waktu respons (rata-rata)
47.94s
Total Biaya
$0.236
Peringkat
#61
Total token output
169,329
Waktu respons (rata-rata)
39.11s
Total Biaya
T/A
Model yang direkomendasikan Seed-2.0-Lite (medium)

It has the best overall balance of score, reliability, cost, and response time in this comparison.

Perbandingan terperinci

Metrik Seed-2.0-Lite Seed-2.0-Lite medium Rilis: 2026-02-14 Qwen3.8 27B Qwen3.8 27B low Rilis: 2026-08-14
Skor 7.8 7.9
Peringkat #69 #61
Keandalan 10.0 10.0
Konsistensi 8.6 10.0
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 69.7% 68.2%
Tes tidak stabil 4 0
Total Run 66 66
Biaya per hasil 1.809 T/A
Total Biaya $0.236 T/A
Harga input $0.250 / 1M T/A
Harga output $2.000 / 1M T/A
Total token input 129,906 99,705
Token output 12,530 1,545
Token penalaran 88,766 167,784
Waktu respons (rata-rata) 47.94s 39.11s
Waktu respons (maks) 254.92s 376.04s
Waktu respons (total) 1054.57s 860.51s
Parameter ~200B total (~20B aktif) 27.3B
Ketersediaan Tertutup Bobot tersedia

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#69 Seed-2.0-Lite

medium
Biaya
$0.005
Waktu
86.7s
Token
2,354 tok

#61 Qwen3.8 27B

low
Biaya
T/A
Waktu
92.8s
Token
6,902 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Seed-2.0-Lite 8.0 9.8 66.7% 0 156.74s 8,247 458 31,890
Qwen3.8 27B 7.7 10.0 66.7% 0 140.92s 8,127 585 86,694

Perbandingan Cepat

Ganti Pasangan Perbandingan

Claude Opus 4.8lowvsSeed-2.0-LitemediumSeed-2.0-LitemediumvsQwen3.7 FlashhighQwen3.8 27BlowvsStep 3.7 FlashmediumGPT-5.2 ChatnonevsQwen3.8 27BlowKimi K3maxvsQwen3.8 27BlowGPT-5.6 TerrahighvsQwen3.8 27BlowGemini 3.5 Flash LitemediumvsQwen3.8 27BlowQwen3.8 27BlowvsGLM 5.2mediumQwen3.8 27BlowvsInklingmediumGPT-5.6 TerramediumvsQwen3.8 27BlowClaude Sonnet 4.6mediumvsQwen3.8 27BlowQwen3.8 27BlowvsGLM 5.2high