Navigate
AI BENCHY
Advertise here

Trinity Large Thinking (high) vs North Mini Code

North Mini Code leads on average score with 5.1 vs 4.8. North Mini Code has the lower benchmark cost at $0.000 vs $0.632. North Mini Code is faster at 29.95s vs 77.22s, with pass rates of 39.4% vs 18.2%.

Last updated at: 2026-07-28

Rank
#202
Total Output Tokens
953,758
Response Time (avg)
77.22s
Total Cost
$0.632
Rank
#191
Total Output Tokens
26,786
Response Time (avg)
29.95s
Total Cost
$0.000
Recommended model North Mini Code

It has the best score here (5.1), while responding about 2.6x faster than Trinity Large Thinking (high).

Detailed comparison

Metric Trinity Large Thinking Trinity Large Thinking high Release: 2026-07-28 North Mini Code North Mini Code none Release: 2026-06-18 Free Available
Score 4.8 5.1
Rank #202 #191
Reliability 9.9 8.7
Consistency 7.2 9.9
Tests Correct
Attempt pass rate 39.4% 18.2%
Flaky tests 8 0
Total Runs 66 60
Cost per result 12.640 0.000
Total Cost $0.632 $0.000
Input Price $0.220 / 1M $0.000 / 1M
Output Price $0.850 / 1M $0.000 / 1M
Total Input Tokens 101,767 130,492
Output Tokens 286,218 26,786
Reasoning Tokens 667,540 0
Response Time (avg) 77.22s 29.95s
Response Time (max) 510.21s 159.85s
Response Time (total) 1698.89s 658.82s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#202 Trinity Large Thinking

high
Cost
$0.028
Time
130.1s
Tokens
34,387 tok

#191 North Mini Code

none
Cost
$0.000
Time
266.1s
Tokens
63,551 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Trinity Large Thinking 3.7 4.7 33.3% 2 245.04s 7,204 83,616 266,836
North Mini Code 3.9 10.0 0.0% 0 21.96s 7,119 504 0

Quick Compare

Switch Comparison Pair