Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Kimi K2.5 vs Laguna S 2.1 (medium)

Skor rata-rata hampir imbang di 5.4 vs 5.4. Laguna S 2.1 (medium) memiliki biaya benchmark lebih rendah di $0.053 vs $0.101. Kimi K2.5 lebih cepat di 18.03s vs 58.42s, dengan tingkat keberhasilan 30.3% vs 24.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-04

Peringkat
#243
Total token output
26,643
Waktu respons (rata-rata)
18.03s
Total Biaya
$0.101
Peringkat
#239
Total token output
339,595
Waktu respons (rata-rata)
58.42s
Total Biaya
$0.053
Model yang direkomendasikan Kimi K2.5

It has the best score here (5.4), while responding about 3.2x faster than Laguna S 2.1 (medium).

Perbandingan terperinci

Metrik Kimi K2.5 Kimi K2.5 none Rilis: 2026-01-27 Laguna S 2.1 Laguna S 2.1 medium Rilis: 2026-07-21 Tersedia gratis
Skor 5.4 5.4
Peringkat #243 #239
Keandalan 10.0 10.0
Konsistensi 8.6 8.8
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 30.3% 24.2%
Tes tidak stabil 4 3
Total Run 66 66
Biaya per hasil 2.278 1.451
Total Biaya $0.101 $0.053
Harga input $0.450 / 1M $0.090 / 1M
Harga output $2.250 / 1M $0.180 / 1M
Total token input 89,320 86,647
Token output 26,643 97,298
Token penalaran 0 242,297
Waktu respons (rata-rata) 18.03s 58.42s
Waktu respons (maks) 102.83s 560.31s
Waktu respons (total) 288.42s 1285.26s
Parameter 1T total (32B aktif) 118B total (8B aktif)
Ketersediaan Bobot tersedia Bobot tersedia

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#243 MoonshotAI: Kimi K2.5

none
Biaya
$0.015
Waktu
89.1s
Token
5,421 tok

#239 Laguna S 2.1

medium
Biaya
$0.001
Waktu
4.0s
Token
867 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Kimi K2.5 5.5 10.0 33.3% 0 24.56s 7,311 4,708 0
Laguna S 2.1 5.3 7.2 44.4% 1 158.98s 7,917 93,354 106,718

Perbandingan Cepat

Ganti Pasangan Perbandingan