Navigate
Advertise here

Mercury 2.5 (low) vs MiniMax M2.7 (medium)

Mercury 2.5 (low) leads on average score with 5.1 vs 5.0. Mercury 2.5 (low) has the lower benchmark cost at $0.011 vs $0.208. Mercury 2.5 (low) is faster at 1.35s vs 43.25s, with pass rates of 39.4% vs 45.5%.

Last updated at: 2026-09-08

Compared models

Rank
#265
Total Output Tokens
33,719
Response Time (avg)
1.35s
Total Cost
$0.011
Rank
#270
Total Output Tokens
147,704
Response Time (avg)
43.25s
Total Cost
$0.208
Recommended model Mercury 2.5 (low)

It has the best score here (5.1), while costing about 19.6x less than MiniMax M2.7 (medium).

Detailed comparison

Metric Mercury 2.5 Mercury 2.5 low Release: 2026-09-08 MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18
Score 5.1 5.0
Rank #265 #270
Reliability 9.8 10.0
Consistency 8.1 6.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 39.4% 45.5%
Flaky tests 5 9
Total Runs 66 66
Cost per result 0.177 4.149
Total Cost $0.011 $0.208
Input Price $0.040 / 1M $0.300 / 1M
Output Price $0.150 / 1M $1.200 / 1M
Total Input Tokens 138,020 114,527
Output Tokens 6,083 18,558
Reasoning Tokens 27,636 129,146
Response Time (avg) 1.35s 43.25s
Response Time (max) 7.48s 196.21s
Response Time (total) 29.74s 908.21s
Parameters ~100B 230B total (10B active)
Availability Closed Weights available

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#265 Mercury 2.5

low
Cost
$0.001
Time
2.4s
Tokens
1,236 tok

#270 MiniMax M2.7

medium
Cost
$0.022
Time
22.8s
Tokens
9,250 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Mercury 2.5 5.5 10.0 33.3% 0 972ms 7,909 521 2,269
MiniMax M2.7 5.7 9.1 33.3% 0 101.89s 2,961 1,231 38,841

Quick Compare

Switch Comparison Pair