Navigasi
Advertise here

Qwen3.5-35B-A3B (medium) vs MiMo-V2.5-Pro (medium)

Skor rata-rata hampir imbang di 6.3 vs 6.4. MiMo-V2.5-Pro (medium) memiliki biaya benchmark lebih rendah di $0.332 vs $0.716. MiMo-V2.5-Pro (medium) lebih cepat di 56.72s vs 111.54s, dengan tingkat keberhasilan 65.2% vs 63.8%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-09

Model yang Dibandingkan

Peringkat
#225
Total token output
913,343
Waktu respons (rata-rata)
111.54s
Total Biaya
$0.716
Peringkat
#221
Total token output
223,971
Waktu respons (rata-rata)
56.72s
Total Biaya
$0.332
Model yang direkomendasikan MiMo-V2.5-Pro (medium)

It has the best score here (6.4), while costing about 2.2x less than Qwen3.5-35B-A3B (medium).

Perbandingan terperinci

Metrik Qwen3.5-35B-A3B Qwen3.5-35B-A3B medium Rilis: 2026-02-24 MiMo-V2.5-Pro MiMo-V2.5-Pro medium Rilis: 2026-04-22
Skor 6.3 6.4
Peringkat #225 #221
Keandalan 10.0 10.0
Konsistensi 7.5 7.6
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 65.2% 63.8%
Tes tidak stabil 7 7
Total Run 69 69
Biaya per hasil 10.505 4.559
Total Biaya $0.716 $0.332
Harga input $0.080 / 1M $0.435 / 1M
Harga output $0.750 / 1M $0.870 / 1M
Harga baca cache $0.040 / 1M $0.004 / 1M
Harga tulis cache T/A T/A
Total token input 378,048 253,462
Token output 56,477 40,269
Token penalaran 865,151 185,716
Waktu respons (rata-rata) 111.54s 56.72s
Waktu respons (maks) 950.25s 333.34s
Waktu respons (total) 2565.38s 1304.54s
Parameter 35B total (3B aktif) 1.02T total (42B aktif)
Ketersediaan Sumber terbuka Sumber terbuka

Harga cache berlaku untuk token input. Pembacaan memakai kembali prompt tersimpan; penulisan menyimpannya dan dapat menambah biaya. Token output menggunakan harga output.

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#225 Qwen3.5-35B-A3B

medium
Biaya
$0.009
Waktu
71.4s
Token
8,631 tok

#221 MiMo-V2.5-Pro

medium
Reached the allocated time limit (300 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
300.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Qwen3.5-35B-A3B 5.9 9.3 33.3% 0 206.65s 4,106 23,844 111,462
MiMo-V2.5-Pro 6.2 4.7 66.7% 2 92.07s 6,543 780 51,218

Ganti Pasangan Perbandingan