Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Anthropic: Claude Fable 5 vs Google: Gemini 3.5 Flash

Summary

The average score is effectively tied at 9.2 vs 9.2. Gemini 3.5 Flash (low) has the lower benchmark cost at $0.349 vs $3.165. Gemini 3.5 Flash (low) is faster at 3.27s vs 17.01s, with pass rates of 82.5% vs 90.5%.

Recommended modelGemini 3.5 Flash (low)It has the best score here (9.2), while costing about 9.1x less than Claude Fable 5 (medium).

Last updated at: 2026-07-14

Metric Claude Fable 5 Claude Fable 5 medium Release: 2026-06-10 Gemini 3.5 Flash Gemini 3.5 Flash low Release: 2026-05-19
Score 9.2 9.2
Rank #9 #8
Reliability 10.0 10.0
Consistency 9.6 10.0
Tests Correct
Attempt pass rate 82.5% 90.5%
Flaky tests 1 0
Total Runs 63 63
Cost per result 18.614 1.834
Total Cost $3.165 $0.349
Input Price $10.000 / 1M $1.500 / 1M
Output Price $50.000 / 1M $9.000 / 1M
Total Input Tokens 58,383 36,938
Output Tokens 41,340 2,033
Reasoning Tokens 10,269 30,519
Response Time (avg) 17.01s 3.27s
Response Time (max) 80.80s 9.05s
Response Time (total) 357.17s 68.65s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#9 Claude Fable 5

medium
Cost
$0.606
Time
156.7s
Tokens
12,264 tok

#8 Gemini 3.5 Flash

low
Cost
$0.068
Time
39.1s
Tokens
7,588 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Fable 5 10.0 10.0 100.0% 0 15.59s 10,590 7,383 1,318
Gemini 3.5 Flash 7.8 10.0 66.7% 0 6.71s 8,118 458 13,420

Quick Compare

Switch Comparison Pair