Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

LongCat 2.0 vs Qwen3.6 Max Preview

Skor rata-rata hampir imbang di 6.4 vs 6.4. LongCat 2.0 memiliki biaya benchmark lebih rendah di $0.044 vs $0.228. LongCat 2.0 lebih cepat di 5.17s vs 7.82s, dengan tingkat keberhasilan 40.9% vs 56.1%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-03

Peringkat
#172
Total token output
9,375
Waktu respons (rata-rata)
5.17s
Total Biaya
$0.044
Peringkat
#169
Total token output
19,257
Waktu respons (rata-rata)
7.82s
Total Biaya
$0.228
Model yang direkomendasikan LongCat 2.0

It has the best score here (6.4), while costing about 5.2x less than Qwen3.6 Max Preview.

Perbandingan terperinci

Metrik LongCat 2.0 LongCat 2.0 none Rilis: 2026-07-20 Qwen3.6 Max Preview Qwen3.6 Max Preview none Rilis: 2026-04-20
Skor 6.4 6.4
Peringkat #172 #169
Keandalan 10.0 9.9
Konsistensi 9.3 9.3
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 40.9% 56.1%
Tes tidak stabil 2 2
Total Run 66 66
Biaya per hasil 0.549 2.248
Total Biaya $0.044 $0.228
Harga input $0.300 / 1M $1.028 / 1M
Harga output $1.200 / 1M $6.162 / 1M
Total token input 108,752 106,348
Token output 9,375 19,257
Token penalaran 0 0
Waktu respons (rata-rata) 5.17s 7.82s
Waktu respons (maks) 48.38s 102.62s
Waktu respons (total) 113.83s 172.04s
Parameter 1.6T total (48B aktif) ~1T total (~40B aktif)
Ketersediaan Sumber terbuka Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#172 LongCat 2.0

none
Biaya
$0.027
Waktu
411.7s
Token
22,283 tok

#169 Qwen3.6 Max Preview

none
Biaya
$0.025
Waktu
83.9s
Token
4,066 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
LongCat 2.0 5.5 10.0 33.3% 0 2.85s 7,446 578 0
Qwen3.6 Max Preview 3.8 7.3 22.2% 1 3.12s 7,913 456 0

Perbandingan Cepat

Ganti Pasangan Perbandingan