Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

North Mini Code vs Google: Gemma 4 31B

Summary

North Mini Code vs Gemma 4 31B benchmark comparison: Gemma 4 31B leads on average score with 6.3 vs 5.1. North Mini Code has the lower benchmark cost at $0.000 vs $0.033. North Mini Code is faster at 29.82s vs 56.55s, with pass rates of 19.1% vs 69.8%.

Recommended model: Gemma 4 31B - It has the strongest score in this comparison (6.3) and the best overall balance of cost and response time across all 2 models.

Last updated at: 2026-06-18

Metric North Mini Code North Mini Code none Release: 2026-06-18 Free Available Gemma 4 31B Gemma 4 31B medium Release: 2026-04-02 Free Available
Score 5.1 6.3
Rank #131 #88
Reliability 8.5 10.0
Consistency 9.9 9.4
Tests Correct
Attempt pass rate 19.1% 69.8%
Flaky tests 0 1
Total Runs 57 63
Cost per result 0.000 0.257
Total Cost $0.000 $0.033
Input Price $0.000 / 1M $0.120 / 1M
Output Price $0.000 / 1M $0.350 / 1M
Total Input Tokens 43,264 17,957
Output Tokens 8,278 22,356
Reasoning Tokens 0 65,726
Response Time (avg) 29.82s 56.55s
Response Time (max) 159.85s 437.40s
Response Time (total) 626.26s 1074.41s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#131 North Mini Code

none
Cost
$0.000
Time
266.1s
Tokens
63,551 tok

#88 Gemma 4 31B

medium
Cost
$0.002
Time
45.7s
Tokens
2,696 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.0 10.0 0.0% 0 22.48s 402 4,075 0
Gemma 4 31B 10.0 10.0 100.0% 0 12.89s 816 962 2,046
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.9 10.0 0.0% 0 21.96s 7,119 504 0
Gemma 4 31B 4.3 5.8 22.2% 1 219.76s 5,568 11,098 33,212
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.5 8.7 0.0% 0 159.85s 24,265 2,920 0
Gemma 4 31B 3.0 10.0 0.0% 0 0ms 0 0 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 10.0 10.0 100.0% 0 28.00s 6,819 183 0
Gemma 4 31B 10.0 10.0 100.0% 0 21.11s 8,334 1,822 2,951
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.0 10.0 0.0% 0 14.73s 621 14 0
Gemma 4 31B 7.7 10.0 66.7% 0 38.48s 876 4,349 8,985
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.9 9.6 0.0% 0 34.77s 444 115 0
Gemma 4 31B 10.0 10.0 100.0% 0 9.57s 567 105 888
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 6.5 10.0 50.0% 0 30.68s 597 57 0
Gemma 4 31B 10.0 10.0 100.0% 0 12.76s 777 533 2,035
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.5 10.0 0.0% 0 24.43s 435 353 0
Gemma 4 31B 9.9 10.0 100.0% 0 26.91s 801 1,795 5,595
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 9.5 10.0 100.0% 0 3.64s 2,403 51 0
Gemma 4 31B 3.0 10.0 0.0% 0 0ms 0 0 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
North Mini Code 3.0 10.0 0.0% 0 37.37s 159 6 0
Gemma 4 31B 3.0 10.0 0.0% 0 90.14s 218 1,692 10,014

Quick Compare

Switch Comparison Pair