Navigasi
AI BENCHY
Advertise here

LongCat 2.0 (low) vs MiMo-V2.5-Pro (medium)

MiMo-V2.5-Pro (medium) unggul dalam skor rata-rata dengan 6.8 vs 6.7. MiMo-V2.5-Pro (medium) memiliki biaya benchmark lebih rendah di $0.221 vs $0.412. MiMo-V2.5-Pro (medium) lebih cepat di 48.68s vs 113.61s, dengan tingkat keberhasilan 54.6% vs 63.6%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-04

Peringkat
#157
Total token output
320,594
Waktu respons (rata-rata)
113.61s
Total Biaya
$0.412
Peringkat
#152
Total token output
185,466
Waktu respons (rata-rata)
48.68s
Total Biaya
$0.221
Model yang direkomendasikan MiMo-V2.5-Pro (medium)

It has the best score here (6.8), while costing about 1.9x less than LongCat 2.0 (low).

Perbandingan terperinci

Metrik LongCat 2.0 LongCat 2.0 low Rilis: 2026-07-20 MiMo-V2.5-Pro MiMo-V2.5-Pro medium Rilis: 2026-04-22
Skor 6.7 6.8
Peringkat #157 #152
Keandalan 9.9 10.0
Konsistensi 8.5 7.9
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 54.6% 63.6%
Tes tidak stabil 4 6
Total Run 66 66
Biaya per hasil 4.120 3.790
Total Biaya $0.412 $0.221
Harga input $0.300 / 1M $0.435 / 1M
Harga output $1.200 / 1M $0.870 / 1M
Total token input 90,870 139,892
Token output 5,676 15,569
Token penalaran 314,918 169,897
Waktu respons (rata-rata) 113.61s 48.68s
Waktu respons (maks) 560.39s 333.34s
Waktu respons (total) 2499.46s 1071.02s
Parameter 1.6T total (48B aktif) 1.02T total (42B aktif)
Ketersediaan Sumber terbuka Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#157 LongCat 2.0

low
Biaya
$0.024
Waktu
428.0s
Token
19,557 tok

#152 MiMo-V2.5-Pro

medium
Reached the allocated time limit (300 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
300.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
LongCat 2.0 6.6 4.6 77.8% 2 479.30s 6,544 461 206,627
MiMo-V2.5-Pro 6.2 4.7 66.7% 2 92.07s 6,543 780 51,218

Perbandingan Cepat

Ganti Pasangan Perbandingan