Navigate
Advertise here

Mercury 2.5 (low) vs Inkling

Inkling leads on average score with 5.2 vs 5.1. Mercury 2.5 (low) has the lower benchmark cost at $0.011 vs $0.147. Mercury 2.5 (low) is faster at 1.35s vs 3.47s, with pass rates of 39.4% vs 28.8%.

Last updated at: 2026-09-08

Compared models

Rank
#265
Total Output Tokens
33,719
Response Time (avg)
1.35s
Total Cost
$0.011
Rank
#258
Total Output Tokens
10,554
Response Time (avg)
3.47s
Total Cost
$0.147
Recommended model Mercury 2.5 (low)

Its score stays close to the best score here (5.1 vs 5.2), while costing about 13.9x less than Inkling.

Detailed comparison

Metric Mercury 2.5 Mercury 2.5 low Release: 2026-09-08 Inkling Inkling none Release: 2026-07-18
Score 5.1 5.2
Rank #265 #258
Reliability 9.8 10.0
Consistency 8.1 9.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 39.4% 28.8%
Flaky tests 5 1
Total Runs 66 66
Cost per result 0.177 2.448
Total Cost $0.011 $0.147
Input Price $0.040 / 1M $1.000 / 1M
Output Price $0.150 / 1M $4.050 / 1M
Total Input Tokens 138,020 104,120
Output Tokens 6,083 10,554
Reasoning Tokens 27,636 0
Response Time (avg) 1.35s 3.47s
Response Time (max) 7.48s 48.02s
Response Time (total) 29.74s 76.28s
Parameters ~100B 975B total (41B active)
Availability Closed Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#265 Mercury 2.5

low
Cost
$0.001
Time
2.4s
Tokens
1,236 tok

#258 Thinking Machines: Inkling

none
Provider returned error
Cost
$0.000
Time
5.3s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Mercury 2.5 5.5 10.0 33.3% 0 972ms 7,909 521 2,269
Inkling 4.5 10.0 0.0% 0 1.01s 7,356 436 0

Quick Compare

Switch Comparison Pair