Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Qwen3.6 Plus (medium) vs Grok 4.7 (medium)

Skor rata-rata hampir imbang di 7.3 vs 7.3. Qwen3.6 Plus (medium) memiliki biaya benchmark lebih rendah di $0.391 vs $3.501. Qwen3.6 Plus (medium) lebih cepat di 57.41s vs 117.65s, dengan tingkat keberhasilan 69.6% vs 78.3%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#137
Total token output
246,790
Waktu respons (rata-rata)
57.41s
Total Biaya
$0.391
Peringkat
#134
Total token output
474,745
Waktu respons (rata-rata)
117.65s
Total Biaya
$3.501
Model yang direkomendasikan Qwen3.6 Plus (medium)

It has the best score here (7.3), while costing about 9.0x less than Grok 4.7 (medium).

Perbandingan terperinci

Metrik Qwen3.6 Plus Qwen3.6 Plus medium Rilis: 2026-04-20 Grok 4.7 Grok 4.7 medium Rilis: 2026-09-21
Skor 7.3 7.3
Peringkat #137 #134
Keandalan 10.0 9.8
Konsistensi 9.0 8.3
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 69.6% 78.3%
Tes tidak stabil 3 5
Total Run 69 69
Biaya per hasil 2.605 19.582
Total Biaya $0.391 $3.501
Harga input $0.325 / 1M $2.000 / 1M
Harga output $1.950 / 1M $6.000 / 1M
Total token input 222,601 326,066
Token output 8,185 5,726
Token penalaran 238,605 469,019
Waktu respons (rata-rata) 57.41s 117.65s
Waktu respons (maks) 307.08s 660.89s
Waktu respons (total) 1263.02s 2706.03s
Parameter ~397B total (~17B aktif) ~1.7T total (~170B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#137 Qwen3.6 Plus

medium
Biaya
$0.024
Waktu
219.0s
Token
12,235 tok

#134 SpaceXAI: Grok 4.7

medium
Biaya
$0.047
Waktu
132.3s
Token
9,954 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Qwen3.6 Plus 6.1 7.8 44.4% 1 153.12s 7,098 58 50,586
Grok 4.7 6.6 4.6 77.8% 2 312.17s 10,266 349 172,347

Perbandingan Cepat

Ganti Pasangan Perbandingan