Navigate
AI BENCHY
Compare Charts Methodology
❤️ Made by XCS
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Trinity Large Preview vs OpenAI: GPT-5 Nano

Compare:

Last updated at: 2026-03-06

Metric Trinity Large Preview none Release: 2026-01-27 Free Available OpenAI: GPT-5 Nano medium Release: 2025-08-07
Rank #45 #34
Avg Score 4.2 5.5
Consistency 9.6 7.0
Cost per result 0.000 0.856
Total Cost $0.000 $0.060
Tests Correct
Attempt pass rate 33.3% 60.4%
Flaky tests 1 6
Total Runs 48 48
Output Tokens 1,837 4,386
Reasoning Tokens 0 142,080
Response Time (avg) 3.15s 47.94s
Response Time (max) 8.91s 204.02s
Response Time (total) 50.46s 431.47s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Avg Score vs Response Time (avg)

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 10.0 10.0 0.0% 0 3.59s 587 0
OpenAI: GPT-5 Nano 7.0 10.0 66.7% 0 37.73s 1,107 19,968
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 10.0 10.0 0.0% 0 8.91s 294 0
OpenAI: GPT-5 Nano 10.0 10.0 100.0% 0 65.96s 578 17,984
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 9.9 10.0 100.0% 0 3.26s 186 0
OpenAI: GPT-5 Nano 10.0 1.7 50.0% 2 21.42s 453 10,560
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 4.0 10.0 33.3% 0 877ms 25 0
OpenAI: GPT-5 Nano 4.0 4.4 55.6% 2 204.02s 237 64,448
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 3.0 9.9 0.0% 0 2.86s 124 0
OpenAI: GPT-5 Nano 3.0 10.0 0.0% 0 17.51s 202 4,608
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 3.5 6.7 16.7% 1 1.09s 63 0
OpenAI: GPT-5 Nano 9.0 6.8 83.3% 1 11.90s 382 4,096
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 4.0 10.0 33.3% 0 3.30s 291 0
OpenAI: GPT-5 Nano 4.0 7.2 44.4% 1 19.81s 869 13,440
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Trinity Large Preview 10.0 10.0 100.0% 0 6.67s 267 0
OpenAI: GPT-5 Nano 10.0 10.0 100.0% 0 33.30s 558 6,976

Quick Compare

Switch Comparison Pair