Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Qwen3.5-122B-A10B (medium) vs Grok 4.6 (low)

Skor rata-rata hampir imbang di 7.0 vs 6.9. Grok 4.6 (low) memiliki biaya benchmark lebih rendah di $0.781 vs $1.152. Grok 4.6 (low) lebih cepat di 16.44s vs 66.96s, dengan tingkat keberhasilan 71.0% vs 66.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#158
Total token output
522,190
Waktu respons (rata-rata)
66.96s
Total Biaya
$1.152
Peringkat
#159
Total token output
42,315
Waktu respons (rata-rata)
16.44s
Total Biaya
$0.781
Model yang direkomendasikan Grok 4.6 (low)

It has the best score here (6.9), while responding about 4.1x faster than Qwen3.5-122B-A10B (medium).

Perbandingan terperinci

Metrik Qwen3.5-122B-A10B Qwen3.5-122B-A10B medium Rilis: 2026-02-24 Grok 4.6 Grok 4.6 low Rilis: 2026-08-12
Skor 7.0 6.9
Peringkat #158 #159
Keandalan 10.0 10.0
Konsistensi 8.3 9.6
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 71.0% 66.7%
Tes tidak stabil 5 1
Total Run 69 69
Biaya per hasil 9.196 5.205
Total Biaya $1.152 $0.781
Harga input $0.260 / 1M $2.000 / 1M
Harga output $2.080 / 1M $6.000 / 1M
Total token input 251,695 263,427
Token output 49,415 6,006
Token penalaran 472,775 36,309
Waktu respons (rata-rata) 66.96s 16.44s
Waktu respons (maks) 519.30s 57.97s
Waktu respons (total) 1539.99s 378.14s
Parameter 122B total (10B aktif) ~1.7T total (~170B aktif)
Ketersediaan Sumber terbuka Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#158 Qwen3.5-122B-A10B

medium
Biaya
$0.019
Waktu
48.7s
Token
6,034 tok

#159 SpaceXAI: Grok 4.6

low
Biaya
$0.011
Waktu
26.0s
Token
1,941 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Qwen3.5-122B-A10B 6.0 7.2 55.6% 1 114.48s 7,630 8,057 82,578
Grok 4.6 5.5 7.1 44.4% 1 21.69s 9,579 349 7,740

Perbandingan Cepat

Ganti Pasangan Perbandingan