Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Apodex 1.1 Mini (low) vs Trinity Large Thinking (medium)

The average score is effectively tied at 6.1 vs 6.1. Apodex 1.1 Mini (low) has the lower benchmark cost at $0.000 vs $0.862. Apodex 1.1 Mini (low) is faster at 33.68s vs 84.01s, with pass rates of 63.8% vs 49.3%.

Last updated at: 2026-10-02

Compared models

Rank
#234
Total Output Tokens
508,815
Response Time (avg)
33.68s
Total Cost
$0.000
Rank
#233
Total Output Tokens
1,033,971
Response Time (avg)
84.01s
Total Cost
$0.862
Recommended model Trinity Large Thinking (medium)

It has the strongest score in this comparison (6.1) and the best overall balance of cost and response time across all 2 models.

Detailed comparison

Metric Apodex 1.1 Mini Apodex 1.1 Mini low Release: 2026-10-02 Free Available Trinity Large Thinking Trinity Large Thinking medium Release: 2026-07-28
Score 6.1 6.1
Rank #234 #233
Reliability 10.0 10.0
Consistency 7.9 7.7
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 63.8% 49.3%
Flaky tests 6 7
Total Runs 69 69
Cost per result 0.000 10.772
Total Cost $0.000 $0.862
Input Price $0.000 / 1M $0.250 / 1M
Output Price $0.000 / 1M $0.800 / 1M
Cache Read Price N/A $0.060 / 1M
Cache Write Price N/A N/A
Total Input Tokens 321,894 319,352
Output Tokens 139,334 158,447
Reasoning Tokens 369,481 875,524
Response Time (avg) 33.68s 84.01s
Response Time (max) 163.05s 525.14s
Response Time (total) 774.73s 1932.24s
Parameters 35B total (~3B active) 398B total (13B active)
Availability Open source Weights available

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#234 Apodex 1.1 Mini

low
Cost
$0.000
Time
41.0s
Tokens
9,010 tok

#233 Trinity Large Thinking

medium
Cost
$0.016
Time
75.9s
Tokens
19,401 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Apodex 1.1 Mini 5.1 4.9 44.5% 2 48.81s 7,893 24,750 76,577
Trinity Large Thinking 7.5 10.0 66.7% 0 179.02s 7,248 7,413 351,155

Quick Compare

Switch Comparison Pair