Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Trinity Large Preview vs Mercury 2.5 Preview (low)

Mercury 2.5 Preview (low) leads on average score with 5.0 vs 4.8. Trinity Large Preview has the lower benchmark cost at $0.008 vs $0.011. Mercury 2.5 Preview (low) is faster at 1.31s vs 2.98s, with pass rates of 21.2% vs 47.0%.

Last updated at: 2026-09-22

Compared models

Rank
#306
Total Output Tokens
2,169
Response Time (avg)
2.98s
Total Cost
$0.008
Rank
#290
Total Output Tokens
31,975
Response Time (avg)
1.31s
Total Cost
$0.011
Recommended model Mercury 2.5 Preview (low)

It has the best score here (5.0), while responding about 2.3x faster than Trinity Large Preview.

Detailed comparison

Metric Trinity Large Preview Trinity Large Preview none Release: 2026-01-27 Mercury 2.5 Preview Mercury 2.5 Preview low Release: 2026-09-02
Score 4.8 5.0
Rank #306 #290
Reliability 10.0 10.0
Consistency 8.9 7.0
Attempts 63/66 66/66
Tests Correct
Attempt pass rate 21.2% 47.0%
Flaky tests 2 8
Total Runs 63 66
Cost per result 0.017 0.169
Total Cost $0.008 $0.011
Input Price $0.243 / 1M $0.040 / 1M
Output Price $0.243 / 1M $0.150 / 1M
Total Input Tokens 29,828 133,525
Output Tokens 2,169 7,259
Reasoning Tokens 0 24,716
Response Time (avg) 2.98s 1.31s
Response Time (max) 14.34s 7.29s
Response Time (total) 56.57s 28.77s
Parameters 398B total (13B active) ~100B
Availability Open source Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#306 Trinity Large Preview

none
No endpoints found for arcee-ai/trinity-large-preview:free.
Cost
$0.000
Time
0.0s
Tokens
0 tok

#290 Mercury 2.5 Preview

low
Cost
$0.001
Time
1.5s
Tokens
1,169 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Trinity Large Preview 3.7 7.7 11.1% 1 14.34s 738 397 0
Mercury 2.5 Preview 5.5 10.0 33.3% 0 972ms 7,794 695 2,263

Quick Compare

Switch Comparison Pair