Navigate
AI BENCHY
Advertise here

Anthropic: Claude Sonnet 5 vs xAI: Grok 4.5

The average score is effectively tied at 8.3 vs 8.3. Claude Sonnet 5 (medium) has the lower benchmark cost at $0.922 vs $1.928. Claude Sonnet 5 (medium) is faster at 12.52s vs 61.71s, with pass rates of 80.3% vs 78.8%.

Recommended modelClaude Sonnet 5 (medium)It has the best score here (8.3), while costing about 2.1x less than Grok 4.5 (medium).

Last updated at: 2026-07-25

Metric Claude Sonnet 5 Claude Sonnet 5 medium Release: 2026-06-30 Grok 4.5 Grok 4.5 medium Release: 2026-07-08
Score 8.3 8.3
Rank #29 #28
Reliability 10.0 10.0
Consistency 9.0 8.9
Tests Correct
Attempt pass rate 80.3% 78.8%
Flaky tests 3 3
Total Runs 66 66
Cost per result 5.760 12.049
Total Cost $0.922 $1.928
Input Price $2.000 / 1M $2.000 / 1M
Output Price $10.000 / 1M $6.000 / 1M
Total Input Tokens 145,956 122,146
Output Tokens 52,333 5,514
Reasoning Tokens 10,874 275,053
Response Time (avg) 12.52s 61.71s
Response Time (max) 66.71s 436.38s
Response Time (total) 275.42s 1357.56s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#29 Claude Sonnet 5

medium
Cost
$0.007
Time
6.4s
Tokens
832 tok

#28 xAI: Grok 4.5

medium
Cost
$0.044
Time
59.4s
Tokens
7,512 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 9.0 7.9 88.9% 1 17.28s 10,590 13,153 2,379
Grok 4.5 7.6 7.2 77.8% 1 155.69s 9,579 390 104,634

Quick Compare

Switch Comparison Pair