Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Claude Fable 5.1 (high) vs GPT-5.4 Nano (medium)

Skor rata-rata hampir imbang di 7.9 vs 8.0. GPT-5.4 Nano (medium) memiliki biaya benchmark lebih rendah di $0.187 vs $3.471. Claude Fable 5.1 (high) lebih cepat di 13.30s vs 15.95s, dengan tingkat keberhasilan 85.5% vs 66.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#98
Total token output
40,466
Waktu respons (rata-rata)
13.30s
Total Biaya
$3.471
Peringkat
#93
Total token output
110,910
Waktu respons (rata-rata)
15.95s
Total Biaya
$0.187
Model yang direkomendasikan Claude Fable 5.1 (high)

It has the strongest score in this comparison (7.9) and the best overall balance of cost and response time across all 2 models.

Perbandingan terperinci

Metrik Claude Fable 5.1 Claude Fable 5.1 high Rilis: 2026-09-02 GPT-5.4 Nano GPT-5.4 Nano medium Rilis: 2026-03-17
Skor 7.9 8.0
Peringkat #98 #93
Keandalan 10.0 10.0
Konsistensi 9.7 8.6
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 85.5% 66.7%
Tes tidak stabil 1 4
Total Run 69 69
Biaya per hasil 18.264 1.432
Total Biaya $3.471 $0.187
Harga input $10.000 / 1M $0.200 / 1M
Harga output $50.000 / 1M $1.250 / 1M
Total token input 144,676 237,088
Token output 8,184 8,235
Token penalaran 32,282 102,675
Waktu respons (rata-rata) 13.30s 15.95s
Waktu respons (maks) 40.17s 94.06s
Waktu respons (total) 305.83s 366.93s
Parameter ~9.5T total (~878B aktif) ~120B total (~5B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#98 Claude Fable 5.1

high
Biaya
$0.384
Waktu
102.5s
Token
7,824 tok

#93 GPT-5.4 Nano

medium
Biaya
$0.007
Waktu
24.6s
Token
4,943 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Claude Fable 5.1 8.4 7.4 88.9% 1 13.91s 10,608 520 6,503
GPT-5.4 Nano 6.1 4.7 66.7% 2 19.12s 7,305 516 20,778

Perbandingan Cepat

Ganti Pasangan Perbandingan