Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Anthropic: Claude Sonnet 4.6 vs Anthropic: Claude Sonnet 5

Summary

Claude Sonnet 4.6 vs Claude Sonnet 5 benchmark comparison: Claude Sonnet 4.6 leads on average score with 7.3 vs 5.7. Claude Sonnet 5 has the lower benchmark cost at $0.287 vs $0.316. Claude Sonnet 5 is faster at 4.74s vs 5.04s, with pass rates of 55.6% vs 42.9%.

Recommended model: Claude Sonnet 4.6 - It has the strongest score in this comparison (7.3) and the best overall balance of cost and response time across all 2 models.

Last updated at: 2026-06-30

Metric Claude Sonnet 4.6 Claude Sonnet 4.6 none Release: 2026-02-17 Claude Sonnet 5 Claude Sonnet 5 none Release: 2026-06-30
Score 7.3 5.7
Rank #57 #117
Reliability 10.0 10.0
Consistency 9.7 8.6
Tests Correct
Attempt pass rate 55.6% 42.9%
Flaky tests 1 4
Total Runs 63 63
Cost per result 2.870 4.098
Total Cost $0.316 $0.287
Input Price $3.000 / 1M $2.000 / 1M
Output Price $15.000 / 1M $10.000 / 1M
Total Input Tokens 57,886 76,797
Output Tokens 9,465 13,325
Reasoning Tokens 0 0
Response Time (avg) 5.04s 4.74s
Response Time (max) 23.84s 29.46s
Response Time (total) 70.60s 99.46s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#57 Claude Sonnet 4.6

none
Cost
$0.038
Time
27.3s
Tokens
2,598 tok

#117 Claude Sonnet 5

none
Cost
$0.061
Time
53.7s
Tokens
6,172 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 4.8 10.0 25.0% 0 2.94s 636 1,214 0
Claude Sonnet 5 5.3 10.0 25.0% 0 3.60s 834 1,813 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 5.5 10.0 33.3% 0 5.19s 8,522 2,127 0
Claude Sonnet 5 4.6 7.9 22.2% 1 3.67s 10,590 1,864 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 9.5 10.0 100.0% 0 23.84s 26,024 3,766 0
Claude Sonnet 5 3.0 10.0 0.0% 0 29.46s 38,775 6,340 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 10.0 10.0 100.0% 0 3.43s 8,574 252 0
Claude Sonnet 5 10.0 10.0 100.0% 0 3.01s 10,503 309 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 7.7 10.0 66.7% 0 3.54s 759 413 0
Claude Sonnet 5 5.3 7.2 44.4% 1 3.28s 975 933 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 6.1 3.1 66.7% 1 2.56s 513 192 0
Claude Sonnet 5 4.7 3.1 33.3% 1 2.81s 708 272 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 6.5 10.0 50.0% 0 1.96s 690 90 0
Claude Sonnet 5 6.4 10.0 50.0% 0 2.58s 909 103 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 7.7 10.0 66.7% 0 2.53s 663 533 0
Claude Sonnet 5 6.0 7.4 55.6% 1 3.22s 894 778 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 10.0 10.0 100.0% 0 4.11s 11,301 447 0
Claude Sonnet 5 10.0 10.0 100.0% 0 6.80s 12,351 522 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 4.6 3.0 10.0 0.0% 0 4.67s 204 431 0
Claude Sonnet 5 3.0 10.0 0.0% 0 4.31s 258 391 0

Quick Compare

Switch Comparison Pair