Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

North Mini Code (medium) vs Mistral Large 4

Mistral Large 4 leads on average score with 5.5 vs 5.4. North Mini Code (medium) has the lower benchmark cost at $0.000 vs $0.353. Mistral Large 4 is faster at 28.46s vs 140.39s, with pass rates of 42.0% vs 30.4%.

Last updated at: 2026-10-06

Compared models

Rank
#285
Total Output Tokens
1,818,271
Response Time (avg)
140.39s
Total Cost
$0.000
Rank
#274
Total Output Tokens
84,218
Response Time (avg)
28.46s
Total Cost
$0.353
Recommended model Mistral Large 4

It has the best score here (5.5), while responding about 4.9x faster than North Mini Code (medium).

Detailed comparison

Metric North Mini Code North Mini Code medium Release: 2026-06-18 Free Available Mistral Large 4 Mistral Large 4 none Release: 2026-10-06
Score 5.4 5.5
Rank #285 #274
Reliability 8.9 9.9
Consistency 8.6 8.8
Attempts 61/69 69/69
Tests Correct
Attempt pass rate 42.0% 30.4%
Flaky tests 4 4
Total Runs 61 69
Cost per result 0.000 7.044
Total Cost $0.000 $0.353
Input Price $0.000 / 1M $0.680 / 1M
Output Price $0.000 / 1M $2.090 / 1M
Cache Read Price N/A $0.070 / 1M
Cache Write Price N/A N/A
Total Input Tokens 166,347 259,048
Output Tokens 423,884 84,218
Reasoning Tokens 1,394,387 0
Response Time (avg) 140.39s 28.46s
Response Time (max) 786.72s 385.10s
Response Time (total) 3229.07s 654.47s
Parameters 30B total (3B active) 1.05T total (49B active)
Availability Open source Closed

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#285 North Mini Code

medium
Cost
$0.000
Time
51.8s
Tokens
12,460 tok

#274 Mistral Large 4

none
Cost
$0.004
Time
26.7s
Tokens
1,923 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 4.5 4.9 33.3% 2 320.43s 7,119 219,891 561,569
Mistral Large 4 5.4 7.8 22.2% 1 153.02s 7,431 62,517 0

Quick Compare

Switch Comparison Pair

North Mini CodemediumFree AvailablevsHy4 previewlowMistral Large 4nonevsNemotron 3 SupermediumFree AvailableNorth Mini CodemediumFree AvailablevsQwen3.8 27BnoneFree AvailableNorth Mini CodemediumFree AvailablevsMiMo-V2.6-FlashnoneNorth Mini CodemediumFree AvailablevsEmber-1noneSeed-2.0-CodenonevsNorth Mini CodemediumFree AvailableNorth Mini CodemediumFree AvailablevsGPT-5.6 LunanoneNorth Mini CodemediumFree AvailablevsGemma 4 26B A4BnoneFree AvailableNorth Mini CodemediumFree AvailablevsGranite 4.2 8BlowNorth Mini CodemediumFree AvailablevsGPT-6 LunanoneNorth Mini CodemediumFree AvailablevsGemini 2.5 FlashnoneNorth Mini CodemediumFree AvailablevsLing 3.1 Flashnone