Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Claude Fable 5.1 (medium) vs DeepSeek V4 Pro 0423

DeepSeek V4 Pro 0423 leads on average score with 7.4 vs 7.3. DeepSeek V4 Pro 0423 has the lower benchmark cost at $0.311 vs $3.197. Claude Fable 5.1 (medium) is faster at 11.87s vs 14.05s, with pass rates of 78.3% vs 49.3%.

Last updated at: 2026-10-01

Compared models

Rank
#133
Total Output Tokens
32,506
Response Time (avg)
11.87s
Total Cost
$3.197
Rank
#129
Total Output Tokens
46,004
Response Time (avg)
14.05s
Total Cost
$0.311
Recommended model DeepSeek V4 Pro 0423

It has the best score here (7.4), while costing about 10.3x less than Claude Fable 5.1 (medium).

Detailed comparison

Metric Claude Fable 5.1 Claude Fable 5.1 medium Release: 2026-09-02 DeepSeek V4 Pro 0423 DeepSeek V4 Pro 0423 none Release: 2026-04-24
Score 7.3 7.4
Rank #133 #129
Reliability 10.0 10.0
Consistency 7.9 8.7
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 78.3% 49.3%
Flaky tests 6 4
Total Runs 69 69
Cost per result 21.312 2.341
Total Cost $3.197 $0.311
Input Price $10.000 / 1M $0.783 / 1M
Output Price $50.000 / 1M $1.566 / 1M
Total Input Tokens 157,142 304,186
Output Tokens 8,149 46,004
Reasoning Tokens 24,357 0
Response Time (avg) 11.87s 14.05s
Response Time (max) 36.59s 119.44s
Response Time (total) 272.91s 323.17s
Parameters ~9.5T total (~878B active) 1.6T total (49B active)
Availability Closed Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#133 Claude Fable 5.1

medium
Cost
$0.244
Time
65.9s
Tokens
5,025 tok

#129 DeepSeek V4 Pro 0423

none
Reached the allocated time limit (300 seconds) without receiving showcase output.
Cost
$0.000
Time
300.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Fable 5.1 7.0 5.0 77.8% 2 10.99s 10,608 525 4,376
DeepSeek V4 Pro 0423 5.6 10.0 33.3% 0 13.38s 7,275 5,500 0

Quick Compare

Switch Comparison Pair