Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Anthropic: Claude Sonnet 5 vs MiniMax: MiniMax M2.7

Summary

Claude Sonnet 5 vs MiniMax M2.7 benchmark comparison: Claude Sonnet 5 leads on average score with 5.7 vs 5.2. MiniMax M2.7 has the lower benchmark cost at $0.075 vs $0.287. Claude Sonnet 5 is faster at 4.74s vs 38.18s, with pass rates of 42.9% vs 46.0%.

Recommended model: Claude Sonnet 5 - It has the best score here (5.7), while responding about 8.1x faster than MiniMax M2.7.

Last updated at: 2026-06-30

Metric Claude Sonnet 5 Claude Sonnet 5 none Release: 2026-06-30 MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18
Score 5.7 5.2
Rank #117 #130
Reliability 10.0 10.0
Consistency 8.6 6.8
Tests Correct
Attempt pass rate 42.9% 46.0%
Flaky tests 4 8
Total Runs 63 63
Cost per result 4.098 2.494
Total Cost $0.287 $0.075
Input Price $2.000 / 1M $0.180 / 1M
Output Price $10.000 / 1M $0.720 / 1M
Total Input Tokens 76,797 34,371
Output Tokens 13,325 8,981
Reasoning Tokens 0 89,812
Response Time (avg) 4.74s 38.18s
Response Time (max) 29.46s 196.21s
Response Time (total) 99.46s 763.60s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#117 Claude Sonnet 5

none
Cost
$0.061
Time
53.7s
Tokens
6,172 tok

#130 MiniMax M2.7

medium
Cost
$0.022
Time
22.8s
Tokens
9,250 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 5.3 10.0 25.0% 0 3.60s 834 1,813 0
MiniMax M2.7 7.9 6.3 83.3% 2 40.32s 654 3,010 17,716
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 4.6 7.9 22.2% 1 3.67s 10,590 1,864 0
MiniMax M2.7 5.7 9.1 33.3% 0 101.89s 2,961 1,231 38,841
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 3.0 10.0 0.0% 0 29.46s 38,775 6,340 0
MiniMax M2.7 4.7 1.6 66.7% 1 41.03s 14,233 369 4,480
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 10.0 10.0 100.0% 0 3.01s 10,503 309 0
MiniMax M2.7 6.3 5.8 66.7% 1 21.95s 7,152 187 5,882
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 5.3 7.2 44.4% 1 3.28s 975 933 0
MiniMax M2.7 3.0 10.0 0.0% 0 19.00s 245 8 2,796
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 4.7 3.1 33.3% 1 2.81s 708 272 0
MiniMax M2.7 3.9 2.5 33.3% 1 38.70s 486 92 5,204
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 6.4 10.0 50.0% 0 2.58s 909 103 0
MiniMax M2.7 3.8 5.8 33.3% 1 12.80s 687 350 2,600
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 6.0 7.4 55.6% 1 3.22s 894 778 0
MiniMax M2.7 5.9 7.2 55.6% 1 24.87s 675 362 7,840
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 10.0 10.0 100.0% 0 6.80s 12,351 522 0
MiniMax M2.7 4.7 1.6 66.7% 1 12.05s 7,067 304 1,001
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Sonnet 5 3.0 10.0 0.0% 0 4.31s 258 391 0
MiniMax M2.7 3.0 10.0 0.0% 0 22.77s 211 3,068 3,452

Quick Compare

Switch Comparison Pair