Navigate
Advertise here

North Mini Code vs Mercury 2.5 (low)

The average score is effectively tied at 5.1 vs 5.1. North Mini Code has the lower benchmark cost at $0.000 vs $0.011. Mercury 2.5 (low) is faster at 1.35s vs 29.95s, with pass rates of 18.2% vs 39.4%.

Last updated at: 2026-09-08

Compared models

Rank
#269
Total Output Tokens
26,786
Response Time (avg)
29.95s
Total Cost
$0.000
Rank
#265
Total Output Tokens
33,719
Response Time (avg)
1.35s
Total Cost
$0.011
Recommended model Mercury 2.5 (low)

It has the best score here (5.1), while responding about 22.2x faster than North Mini Code.

Detailed comparison

Metric North Mini Code North Mini Code none Release: 2026-06-18 Free Available Mercury 2.5 Mercury 2.5 low Release: 2026-09-08
Score 5.1 5.1
Rank #269 #265
Reliability 8.7 9.8
Consistency 9.9 8.1
Attempts 60/66 66/66
Tests Correct
Attempt pass rate 18.2% 39.4%
Flaky tests 0 5
Total Runs 60 66
Cost per result 0.000 0.177
Total Cost $0.000 $0.011
Input Price $0.000 / 1M $0.040 / 1M
Output Price $0.000 / 1M $0.150 / 1M
Total Input Tokens 130,501 138,020
Output Tokens 26,786 6,083
Reasoning Tokens 0 27,636
Response Time (avg) 29.95s 1.35s
Response Time (max) 159.85s 7.48s
Response Time (total) 658.83s 29.74s
Parameters 30B total (3B active) ~100B
Availability Open source Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#269 North Mini Code

none
Cost
$0.000
Time
266.1s
Tokens
63,551 tok

#265 Mercury 2.5

low
Cost
$0.001
Time
2.4s
Tokens
1,236 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.9 10.0 0.0% 0 21.96s 7,119 504 0
Mercury 2.5 5.5 10.0 33.3% 0 972ms 7,909 521 2,269

Quick Compare

Switch Comparison Pair