Navigate
AI BENCHY
Advertise here

Trinity Large Thinking (low) vs GPT-5 Nano (medium)

The average score is effectively tied at 6.0 vs 6.1. GPT-5 Nano (medium) has the lower benchmark cost at $0.114 vs $0.656. GPT-5 Nano (medium) is faster at 54.87s vs 96.78s, with pass rates of 48.5% vs 56.1%.

Last updated at: 2026-07-28

Rank
#145
Total Output Tokens
815,220
Response Time (avg)
96.78s
Total Cost
$0.656
Rank
#143
Total Output Tokens
273,098
Response Time (avg)
54.87s
Total Cost
$0.114
Recommended model GPT-5 Nano (medium)

It has the best score here (6.1), while costing about 5.8x less than Trinity Large Thinking (low).

Detailed comparison

Metric Trinity Large Thinking Trinity Large Thinking low Release: 2026-07-28 GPT-5 Nano GPT-5 Nano medium Release: 2025-08-07
Score 6.0 6.1
Rank #145 #143
Reliability 9.7 10.0
Consistency 7.9 7.0
Tests Correct
Attempt pass rate 48.5% 56.1%
Flaky tests 6 8
Total Runs 66 66
Cost per result 8.199 1.267
Total Cost $0.656 $0.114
Input Price $0.220 / 1M $0.050 / 1M
Output Price $0.850 / 1M $0.400 / 1M
Total Input Tokens 122,845 94,935
Output Tokens 117,704 12,042
Reasoning Tokens 697,516 261,056
Response Time (avg) 96.78s 54.87s
Response Time (max) 540.96s 227.89s
Response Time (total) 2129.13s 822.99s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#145 Trinity Large Thinking

low
Cost
$0.021
Time
173.2s
Tokens
24,586 tok

#143 GPT-5 Nano

medium
Cost
$0.006
Time
108.5s
Tokens
13,209 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Trinity Large Thinking 5.3 10.0 33.3% 0 344.45s 5,441 27,039 217,241
GPT-5 Nano 7.0 7.7 55.6% 1 41.62s 7,305 740 41,152

Quick Compare

Switch Comparison Pair