Navigate
Advertise here

Claude Fable 5.1 (medium) vs Gemma 4 26B A4B (medium)

The average score is effectively tied at 7.3 vs 7.3. Gemma 4 26B A4B (medium) has the lower benchmark cost at $0.101 vs $3.197. Claude Fable 5.1 (medium) is faster at 11.87s vs 105.03s, with pass rates of 78.3% vs 71.0%.

Last updated at: 2026-10-02

Compared models

Rank
#133
Total Output Tokens
32,506
Response Time (avg)
11.87s
Total Cost
$3.197
Rank
#131
Total Output Tokens
269,351
Response Time (avg)
105.03s
Total Cost
$0.101
Recommended model Gemma 4 26B A4B (medium)

It has the best score here (7.3), while costing about 31.7x less than Claude Fable 5.1 (medium).

Detailed comparison

Metric Claude Fable 5.1 Claude Fable 5.1 medium Release: 2026-09-02 Gemma 4 26B A4B Gemma 4 26B A4B medium Release: 2026-04-03 Free Available
Score 7.3 7.3
Rank #133 #131
Reliability 10.0 10.0
Consistency 7.9 9.6
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 78.3% 71.0%
Flaky tests 6 1
Total Runs 69 69
Cost per result 21.312 0.679
Total Cost $3.197 $0.101
Input Price $10.000 / 1M $0.090 / 1M
Output Price $50.000 / 1M $0.300 / 1M
Cache Read Price $0.250 / 1M $0.050 / 1M
Cache Write Price $12.500 / 1M N/A
Total Input Tokens 157,142 224,367
Output Tokens 8,149 27,095
Reasoning Tokens 24,357 242,256
Response Time (avg) 11.87s 105.03s
Response Time (max) 36.59s 912.19s
Response Time (total) 272.91s 2310.63s
Parameters ~9.5T total (~878B active) 25.2B total (3.8B active)
Availability Closed Open source

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#133 Claude Fable 5.1

medium
Cost
$0.244
Time
65.9s
Tokens
5,025 tok

#131 Gemma 4 26B A4B

medium
Reached the allocated time limit (300 seconds) without receiving showcase output.
Cost
$0.000
Time
300.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Fable 5.1 7.0 5.0 77.8% 2 10.99s 10,608 525 4,376
Gemma 4 26B A4B 2.9 10.0 0.0% 0 272.54s 5,062 14,838 44,567

Quick Compare

Switch Comparison Pair