Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

gpt-oss-120b vs Qwen3 Coder Next (medium)

Qwen3 Coder Next (medium) unggul dalam skor rata-rata dengan 4.6 vs 3.7. gpt-oss-120b memiliki biaya benchmark lebih rendah di $0.033 vs $0.034. Qwen3 Coder Next (medium) lebih cepat di 9.07s vs 21.61s, dengan tingkat keberhasilan 33.3% vs 25.8%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-18

Model yang Dibandingkan

Peringkat
#325
Total token output
51,664
Waktu respons (rata-rata)
21.61s
Total Biaya
$0.033
Peringkat
#300
Total token output
19,069
Waktu respons (rata-rata)
9.07s
Total Biaya
$0.034
Model yang direkomendasikan Qwen3 Coder Next (medium)

It has the best score here (4.6), while responding about 2.4x faster than gpt-oss-120b.

Perbandingan terperinci

Metrik gpt-oss-120b gpt-oss-120b none Rilis: 2025-08-05 Tersedia gratis Qwen3 Coder Next Qwen3 Coder Next medium Rilis: 2026-02-03
Skor 3.7 4.6
Peringkat #325 #300
Keandalan 10.0 10.0
Konsistensi 7.8 8.6
Percobaan 57/66 66/66
Tes benar
Tingkat lulus per percobaan 33.3% 25.8%
Tes tidak stabil 2 4
Total Run 57 66
Biaya per hasil 0.168 1.058
Total Biaya $0.033 $0.034
Harga input $0.150 / 1M $0.120 / 1M
Harga output $0.600 / 1M $0.800 / 1M
Total token input 9,081 148,203
Token output 51,664 19,069
Token penalaran 0 0
Waktu respons (rata-rata) 21.61s 9.07s
Waktu respons (maks) 113.71s 81.80s
Waktu respons (total) 345.79s 154.19s
Parameter 117B total (5.1B aktif) 80B total (3B aktif)
Ketersediaan Sumber terbuka Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#325 gpt-oss-120b

none
Belum ada hasil showcase yang dihasilkan untuk model ini.
Biaya
T/A
Waktu
-
Token
0 tok

#300 Qwen3 Coder Next

medium
Reached the allocated time limit (300 seconds) without receiving showcase output.
Biaya
$0.000
Waktu
300.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
gpt-oss-120b 1.5 4.0 22.2% 1 9.57s 901 3,232 0
Qwen3 Coder Next 3.7 7.2 22.2% 1 924ms 7,185 336 0

Perbandingan Cepat

Ganti Pasangan Perbandingan