Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Qwen3.7 Max vs Grok 4.6 (low)

Qwen3.7 Max unggul dalam skor rata-rata dengan 7.4 vs 7.4. Qwen3.7 Max memiliki biaya benchmark lebih rendah di $0.197 vs $0.449. Qwen3.7 Max lebih cepat di 4.56s vs 14.55s, dengan tingkat keberhasilan 68.2% vs 69.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-24

Model yang Dibandingkan

Peringkat
#129
Total token output
12,446
Waktu respons (rata-rata)
4.56s
Total Biaya
$0.197
Peringkat
#135
Total token output
38,048
Waktu respons (rata-rata)
14.55s
Total Biaya
$0.449
Model yang direkomendasikan Qwen3.7 Max

It has the best score here (7.4), while costing about 2.3x less than Grok 4.6 (low).

Perbandingan terperinci

Metrik Qwen3.7 Max Qwen3.7 Max none Rilis: 2026-05-22 Grok 4.6 Grok 4.6 low Rilis: 2026-08-12
Skor 7.4 7.4
Peringkat #129 #135
Keandalan 9.9 10.0
Konsistensi 10.0 9.6
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 68.2% 69.7%
Tes tidak stabil 0 1
Total Run 66 66
Biaya per hasil 1.581 2.989
Total Biaya $0.197 $0.449
Harga input $1.475 / 1M $2.000 / 1M
Harga output $4.425 / 1M $6.000 / 1M
Total token input 95,992 109,960
Token output 12,446 5,124
Token penalaran 0 32,924
Waktu respons (rata-rata) 4.56s 14.55s
Waktu respons (maks) 72.30s 26.46s
Waktu respons (total) 100.30s 320.17s
Parameter ~1T total (~40B aktif) ~1.7T total (~170B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#129 Qwen3.7 Max

none
Biaya
$0.046
Waktu
195.0s
Token
12,171 tok

#135 SpaceXAI: Grok 4.6

low
Biaya
$0.011
Waktu
26.0s
Token
1,941 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Qwen3.7 Max 5.5 10.0 33.3% 0 1.35s 7,911 582 0
Grok 4.6 5.5 7.1 44.4% 1 21.69s 9,579 349 7,740

Perbandingan Cepat

Ganti Pasangan Perbandingan