Navigasi
Advertise here

LongCat 2.0 (low) vs Inkling Small (medium)

LongCat 2.0 (low) unggul dalam skor rata-rata dengan 6.7 vs 6.6. Inkling Small (medium) memiliki biaya benchmark lebih rendah di $0.113 vs $0.412. Inkling Small (medium) lebih cepat di 6.30s vs 113.61s, dengan tingkat keberhasilan 54.6% vs 53.0%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-10

Model yang Dibandingkan

Peringkat
#166
Total token output
320,594
Waktu respons (rata-rata)
113.61s
Total Biaya
$0.412
Peringkat
#174
Total token output
56,269
Waktu respons (rata-rata)
6.30s
Total Biaya
$0.113
Model yang direkomendasikan Inkling Small (medium)

Its score stays close to the best score here (6.6 vs 6.7), while costing about 3.7x less than LongCat 2.0 (low).

Perbandingan terperinci

Metrik LongCat 2.0 LongCat 2.0 low Rilis: 2026-07-20 Inkling Small Inkling Small medium Rilis: 2026-08-01
Skor 6.7 6.6
Peringkat #166 #174
Keandalan 9.9 10.0
Konsistensi 8.5 9.0
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 54.6% 53.0%
Tes tidak stabil 4 3
Total Run 66 66
Biaya per hasil 4.120 1.177
Total Biaya $0.412 $0.113
Harga input $0.300 / 1M $0.450 / 1M
Harga output $1.200 / 1M $1.200 / 1M
Total token input 90,870 100,255
Token output 5,676 6,422
Token penalaran 314,918 49,847
Waktu respons (rata-rata) 113.61s 6.30s
Waktu respons (maks) 560.39s 17.26s
Waktu respons (total) 2499.46s 138.60s
Parameter 1.6T total (48B aktif) 276B total (12B aktif)
Ketersediaan Sumber terbuka Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#166 LongCat 2.0

low
Biaya
$0.024
Waktu
428.0s
Token
19,557 tok

#174 Thinking Machines: Inkling Small

medium
Biaya
$0.003
Waktu
13.4s
Token
1,753 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
LongCat 2.0 6.6 4.6 77.8% 2 479.30s 6,544 461 206,627
Inkling Small 7.8 10.0 66.7% 0 10.49s 7,374 434 13,783

Perbandingan Cepat

Ganti Pasangan Perbandingan