Navigasi
AI BENCHY
Advertise here

LongCat 2.0 (high) vs Qwen3.6 35B A3B (medium)

LongCat 2.0 (high) unggul dalam skor rata-rata dengan 6.7 vs 6.6. LongCat 2.0 (high) memiliki biaya benchmark lebih rendah di $0.492 vs $0.672. Qwen3.6 35B A3B (medium) lebih cepat di 58.49s vs 153.29s, dengan tingkat keberhasilan 51.5% vs 57.6%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-04

Peringkat
#160
Total token output
386,642
Waktu respons (rata-rata)
153.29s
Total Biaya
$0.492
Peringkat
#167
Total token output
743,292
Waktu respons (rata-rata)
58.49s
Total Biaya
$0.672
Model yang direkomendasikan Qwen3.6 35B A3B (medium)

Its score stays close to the best score here (6.6 vs 6.7), while responding about 2.6x faster than LongCat 2.0 (high).

Perbandingan terperinci

Metrik LongCat 2.0 LongCat 2.0 high Rilis: 2026-07-20 Qwen3.6 35B A3B Qwen3.6 35B A3B medium Rilis: 2026-04-20
Skor 6.7 6.6
Peringkat #160 #167
Keandalan 10.0 10.0
Konsistensi 8.4 9.2
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 51.5% 57.6%
Tes tidak stabil 4 2
Total Run 66 66
Biaya per hasil 5.467 6.212
Total Biaya $0.492 $0.672
Harga input $0.300 / 1M $0.100 / 1M
Harga output $1.200 / 1M $0.900 / 1M
Total token input 93,411 85,148
Token output 30,470 61,736
Token penalaran 356,172 681,556
Waktu respons (rata-rata) 153.29s 58.49s
Waktu respons (maks) 941.66s 817.57s
Waktu respons (total) 3372.36s 1169.83s
Parameter 1.6T total (48B aktif) 35B total (3B aktif)
Ketersediaan Sumber terbuka Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#160 LongCat 2.0

high
Reached the allocated time limit (600 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
600.0s
Token
0 tok

#167 Qwen3.6 35B A3B

medium
Reached the allocated time limit (300 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
300.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
LongCat 2.0 7.8 9.3 66.7% 0 505.33s 6,913 244 192,377
Qwen3.6 35B A3B 7.7 10.0 66.7% 0 50.55s 5,051 7,929 37,223

Perbandingan Cepat

Ganti Pasangan Perbandingan