Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

LongCat 2.0 vs Qwen3.5-35B-A3B (medium)

Skor rata-rata hampir imbang di 6.4 vs 6.3. LongCat 2.0 memiliki biaya benchmark lebih rendah di $0.104 vs $1.249. LongCat 2.0 lebih cepat di 8.50s vs 111.54s, dengan tingkat keberhasilan 42.0% vs 65.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#211
Total token output
17,473
Waktu respons (rata-rata)
8.50s
Total Biaya
$0.104
Peringkat
#212
Total token output
921,628
Waktu respons (rata-rata)
111.54s
Total Biaya
$1.249
Model yang direkomendasikan LongCat 2.0

It has the best score here (6.4), while costing about 12.0x less than Qwen3.5-35B-A3B (medium).

Perbandingan terperinci

Metrik LongCat 2.0 LongCat 2.0 none Rilis: 2026-07-20 Qwen3.5-35B-A3B Qwen3.5-35B-A3B medium Rilis: 2026-02-24
Skor 6.4 6.3
Peringkat #211 #212
Keandalan 10.0 10.0
Konsistensi 9.0 7.5
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 42.0% 65.2%
Tes tidak stabil 3 7
Total Run 69 69
Biaya per hasil 1.298 10.505
Total Biaya $0.104 $1.249
Harga input $0.300 / 1M $0.163 / 1M
Harga output $1.200 / 1M $1.300 / 1M
Total token input 276,199 378,048
Token output 17,473 56,477
Token penalaran 0 865,151
Waktu respons (rata-rata) 8.50s 111.54s
Waktu respons (maks) 81.77s 950.25s
Waktu respons (total) 195.60s 2565.38s
Parameter 1.6T total (48B aktif) 35B total (3B aktif)
Ketersediaan Sumber terbuka Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#211 LongCat 2.0

none
Biaya
$0.027
Waktu
411.7s
Token
22,283 tok

#212 Qwen3.5-35B-A3B

medium
Biaya
$0.009
Waktu
71.4s
Token
8,631 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
LongCat 2.0 5.5 10.0 33.3% 0 2.85s 7,446 578 0
Qwen3.5-35B-A3B 5.9 9.3 33.3% 0 206.65s 4,106 23,844 111,462

Perbandingan Cepat

Ganti Pasangan Perbandingan