Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Claude Fable 5.1 (medium) vs GPT-5.2 (medium)

GPT-5.2 (medium) unggul dalam skor rata-rata dengan 8.4 vs 8.3. GPT-5.2 (medium) memiliki biaya benchmark lebih rendah di $1.182 vs $2.909. Claude Fable 5.1 (medium) lebih cepat di 11.67s vs 28.51s, dengan tingkat keberhasilan 81.8% vs 72.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-02

Peringkat
#48
Total token output
31,152
Waktu respons (rata-rata)
11.67s
Total Biaya
$2.909
Peringkat
#43
Total token output
71,252
Waktu respons (rata-rata)
28.51s
Total Biaya
$1.182
Model yang direkomendasikan GPT-5.2 (medium)

It has the best score here (8.4), while costing about 2.5x less than Claude Fable 5.1 (medium).

Perbandingan terperinci

Metrik Claude Fable 5.1 Claude Fable 5.1 medium Rilis: 2026-09-02 GPT-5.2 GPT-5.2 medium Rilis: 2025-12-11
Skor 8.3 8.4
Peringkat #48 #43
Keandalan 10.0 10.0
Konsistensi 7.8 8.5
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 81.8% 72.7%
Tes tidak stabil 6 4
Total Run 66 66
Biaya per hasil 19.392 8.438
Total Biaya $2.909 $1.182
Harga input $10.000 / 1M $1.750 / 1M
Harga output $50.000 / 1M $14.000 / 1M
Total token input 135,112 105,013
Token output 7,594 9,914
Token penalaran 23,558 61,338
Waktu respons (rata-rata) 11.67s 28.51s
Waktu respons (maks) 36.59s 116.86s
Waktu respons (total) 256.73s 456.13s
Parameter ~9.5T total (~878B aktif) ~1.2T total (~80B aktif)
Ketersediaan Tertutup Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#48 Claude Fable 5.1

medium
Biaya
$0.244
Waktu
65.9s
Token
5,025 tok

#43 GPT-5.2

medium
Biaya
$0.047
Waktu
49.2s
Token
3,396 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Claude Fable 5.1 7.0 5.0 77.8% 2 10.99s 10,608 525 4,376
GPT-5.2 10.0 10.0 100.0% 0 22.73s 7,302 511 11,912

Perbandingan Cepat

Ganti Pasangan Perbandingan