Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Claude Fable 5 (medium) vs GPT-5.4 (medium)

Claude Fable 5 (medium) leads on average score with 8.6 vs 8.5. GPT-5.4 (medium) has the lower benchmark cost at $1.560 vs $3.004. Claude Fable 5 (medium) is faster at 15.01s vs 22.85s, with pass rates of 78.8% vs 77.3%.

Last updated at: 2026-09-02

Rank
#40
Total Output Tokens
42,146
Response Time (avg)
15.01s
Total Cost
$3.004
Rank
#43
Total Output Tokens
90,466
Response Time (avg)
22.85s
Total Cost
$1.560
Recommended model GPT-5.4 (medium)

Its score stays close to the best score here (8.5 vs 8.6), while costing about 1.9x less than Claude Fable 5 (medium).

Detailed comparison

Metric Claude Fable 5 Claude Fable 5 medium Release: 2026-06-10 GPT-5.4 GPT-5.4 medium Release: 2026-03-05
Score 8.6 8.5
Rank #40 #43
Reliability 10.0 10.0
Consistency 9.6 8.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 78.8% 77.3%
Flaky tests 1 4
Total Runs 66 66
Cost per result 17.669 10.399
Total Cost $3.004 $1.560
Input Price $10.000 / 1M $2.500 / 1M
Output Price $50.000 / 1M $15.000 / 1M
Total Input Tokens 89,640 81,136
Output Tokens 33,092 6,155
Reasoning Tokens 9,054 84,311
Response Time (avg) 15.01s 22.85s
Response Time (max) 80.80s 100.41s
Response Time (total) 330.15s 502.78s
Parameters ~9.5T total (~878B active) ~1.5T total (~100B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#40 Claude Fable 5

medium
Cost
$0.606
Time
156.7s
Tokens
12,264 tok

#43 GPT-5.4

medium
Cost
$0.214
Time
199.6s
Tokens
14,349 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Fable 5 10.0 10.0 100.0% 0 15.59s 10,590 7,383 1,318
GPT-5.4 8.8 7.8 88.9% 1 44.36s 7,305 433 24,216

Quick Compare

Switch Comparison Pair