Navigate
AI BENCHY
Advertise here

North Mini Code vs Mistral Small 4 (medium)

The average score is effectively tied at 5.1 vs 5.1. North Mini Code has the lower benchmark cost at $0.000 vs $0.096. Mistral Small 4 (medium) is faster at 10.77s vs 29.95s, with pass rates of 18.2% vs 42.4%.

Last updated at: 2026-08-02

Rank
#199
Total Output Tokens
26,786
Response Time (avg)
29.95s
Total Cost
$0.000
Rank
#195
Total Output Tokens
131,824
Response Time (avg)
10.77s
Total Cost
$0.096
Recommended model Mistral Small 4 (medium)

It has the best score here (5.1), while responding about 2.8x faster than North Mini Code.

Detailed comparison

Metric North Mini Code North Mini Code none Release: 2026-06-18 Free Available Mistral Small 4 Mistral Small 4 medium Release: 2026-03-16
Score 5.1 5.1
Rank #199 #195
Reliability 8.7 10.0
Consistency 9.9 7.0
Benchmark coverage 22/22 tests · 60/66 attempts 22/22 tests · 66/66 attempts
Tests Correct
Attempt pass rate 18.2% 42.4%
Flaky tests 0 8
Total Runs 60 66
Cost per result 0.000 1.913
Total Cost $0.000 $0.096
Input Price $0.000 / 1M $0.150 / 1M
Output Price $0.000 / 1M $0.600 / 1M
Total Input Tokens 130,492 140,494
Output Tokens 26,786 39,462
Reasoning Tokens 0 92,362
Response Time (avg) 29.95s 10.77s
Response Time (max) 159.85s 59.15s
Response Time (total) 658.82s 236.94s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#199 North Mini Code

none
Cost
$0.000
Time
266.1s
Tokens
63,551 tok

#195 Mistral Small 4

medium
Cost
$0.006
Time
47.9s
Tokens
9,857 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.9 10.0 0.0% 0 21.96s 7,119 504 0
Mistral Small 4 4.4 5.1 33.3% 2 39.98s 7,636 11,635 54,715

Quick Compare

Switch Comparison Pair