Navigasi
Advertise here

Claude Fable 5.1 (high) vs GPT-5.4 Mini (medium)

Skor rata-rata hampir imbang di 7.9 vs 7.9. GPT-5.4 Mini (medium) memiliki biaya benchmark lebih rendah di $1.008 vs $3.471. Claude Fable 5.1 (high) lebih cepat di 13.30s vs 29.94s, dengan tingkat keberhasilan 85.5% vs 69.6%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#98
Total token output
40,466
Waktu respons (rata-rata)
13.30s
Total Biaya
$3.471
Peringkat
#99
Total token output
184,172
Waktu respons (rata-rata)
29.94s
Total Biaya
$1.008
Model yang direkomendasikan Claude Fable 5.1 (high)

It has the best score here (7.9), while responding about 2.3x faster than GPT-5.4 Mini (medium).

Perbandingan terperinci

Metrik Claude Fable 5.1 Claude Fable 5.1 high Rilis: 2026-09-02 GPT-5.4 Mini GPT-5.4 Mini medium Rilis: 2026-03-17
Skor 7.9 7.9
Peringkat #98 #99
Keandalan 10.0 10.0
Konsistensi 9.7 8.2
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 85.5% 69.6%
Tes tidak stabil 1 5
Total Run 69 69
Biaya per hasil 18.264 7.754
Total Biaya $3.471 $1.008
Harga input $10.000 / 1M $0.750 / 1M
Harga output $50.000 / 1M $4.500 / 1M
Total token input 144,676 238,925
Token output 8,184 7,109
Token penalaran 32,282 177,063
Waktu respons (rata-rata) 13.30s 29.94s
Waktu respons (maks) 40.17s 138.75s
Waktu respons (total) 305.83s 688.57s
Parameter ~9.5T total (~878B aktif) ~400B total (~17B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#98 Claude Fable 5.1

high
Biaya
$0.384
Waktu
102.5s
Token
7,824 tok

#99 GPT-5.4 Mini

medium
Biaya
$0.056
Waktu
95.5s
Token
12,464 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Claude Fable 5.1 8.4 7.4 88.9% 1 13.91s 10,608 520 6,503
GPT-5.4 Mini 8.4 7.4 88.9% 1 57.87s 7,305 467 40,902

Perbandingan Cepat

Ganti Pasangan Perbandingan