Navigate
Advertise here

Mercury 2.5 (high) vs Solar Pro 4 (high)

The average score is effectively tied at 7.1 vs 7.1. Mercury 2.5 (high) has the lower benchmark cost at $0.031 vs $0.046. Mercury 2.5 (high) is faster at 4.32s vs 120.58s, with pass rates of 62.1% vs 62.1%.

Last updated at: 2026-09-08

Compared models

Rank
#134
Total Output Tokens
174,547
Response Time (avg)
4.32s
Total Cost
$0.031
Rank
#133
Total Output Tokens
363,194
Response Time (avg)
120.58s
Total Cost
$0.046
Recommended model Mercury 2.5 (high)

It has the best score here (7.1), while responding about 27.9x faster than Solar Pro 4 (high).

Detailed comparison

Metric Mercury 2.5 Mercury 2.5 high Release: 2026-09-08 Solar Pro 4 Solar Pro 4 high Release: 2026-08-11
Score 7.1 7.1
Rank #134 #133
Reliability 9.7 10.0
Consistency 8.6 8.3
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 62.1% 62.1%
Flaky tests 4 5
Total Runs 66 66
Cost per result 0.259 0.418
Total Cost $0.031 $0.046
Input Price $0.040 / 1M $0.030 / 1M
Output Price $0.150 / 1M $0.120 / 1M
Total Input Tokens 120,068 79,740
Output Tokens 3,402 23,357
Reasoning Tokens 171,145 339,837
Response Time (avg) 4.32s 120.58s
Response Time (max) 23.63s 491.07s
Response Time (total) 95.06s 2652.77s
Parameters ~100B ~100B
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#134 Mercury 2.5

high
Cost
$0.001
Time
3.5s
Tokens
2,447 tok

#133 Solar Pro 4

high
No showcase result has been generated for this model yet.
Cost
N/A
Time
-
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Mercury 2.5 6.4 7.8 44.4% 1 6.58s 7,757 415 44,012
Solar Pro 4 5.7 9.9 33.3% 0 244.73s 7,887 18,577 101,928

Quick Compare

Switch Comparison Pair