- Rank
- #230
- Total Output Tokens
- 26,073
- Response Time (avg)
- 11.40s
- Total Cost
- $3.395
Claude Fable 5.1 (low) vs Trinity Large Thinking (medium)
Claude Fable 5.1 (low) leads on average score with 6.2 vs 6.1. Trinity Large Thinking (medium) has the lower benchmark cost at $0.820 vs $3.395. Claude Fable 5.1 (low) is faster at 11.40s vs 84.01s, with pass rates of 55.1% vs 49.3%.
Compared models
- Rank
- #232
- Total Output Tokens
- 1,033,971
- Response Time (avg)
- 84.01s
- Total Cost
- $0.820
Recommended model
Claude Fable 5.1 (low)
It has the best score here (6.2), while responding about 7.4x faster than Trinity Large Thinking (medium).
Detailed comparison
| Metric | Claude Fable 5.1 Claude Fable 5.1 low | Trinity Large Thinking Trinity Large Thinking medium |
|---|---|---|
| Score | 6.2 | 6.1 |
| Rank | #230 | #232 |
| Reliability | 10.0 | 10.0 |
| Consistency | 8.9 | 7.7 |
| Attempts | 69/69 | 69/69 |
| Tests Correct | ||
| Attempt pass rate | 55.1% | 49.3% |
| Flaky tests | 3 | 7 |
| Total Runs | 69 | 69 |
| Cost per result | 30.862 | 10.772 |
| Total Cost | $3.395 | $0.820 |
| Input Price | $10.000 / 1M | $0.250 / 1M |
| Output Price | $50.000 / 1M | $0.800 / 1M |
| Total Input Tokens | 209,116 | 319,352 |
| Output Tokens | 10,031 | 158,447 |
| Reasoning Tokens | 16,042 | 875,524 |
| Response Time (avg) | 11.40s | 84.01s |
| Response Time (max) | 34.77s | 525.14s |
| Response Time (total) | 262.14s | 1932.24s |
| Parameters | ~9.5T total (~878B active) | 398B total (13B active) |
| Availability | Closed | Weights available |
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#230 Claude Fable 5.1
low- Cost
- $0.126
- Time
- 28.5s
- Tokens
- 2,652 tok
#232 Trinity Large Thinking
medium- Cost
- $0.016
- Time
- 75.9s
- Tokens
- 19,401 tok
Top Models by Score
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
| Agentic | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 2.8 | 1.6 | 33.3% | 1 | 34.77s | 74,534 | 1,076 | 622 | |
| Trinity Large Thinking | 6.1 | 3.1 | 66.7% | 1 | 52.50s | 200,866 | 2,281 | 14,905 |
| Anti-AI Tricks | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 6.5 | 10.0 | 50.0% | 0 | 4.54s | 858 | 666 | 0 | |
| Trinity Large Thinking | 6.8 | 9.8 | 50.0% | 0 | 5.11s | 663 | 1,531 | 6,078 |
| Coding | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 3.4 | 10.0 | 0.0% | 0 | 6.75s | 10,608 | 1,753 | 0 | |
| Trinity Large Thinking | 7.5 | 10.0 | 66.7% | 0 | 179.02s | 7,248 | 7,413 | 351,155 |
| Combined | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 10.0 | 10.0 | 100.0% | 0 | 29.40s | 97,043 | 4,733 | 3,850 | |
| Trinity Large Thinking | 2.9 | 6.0 | 16.7% | 1 | 382.70s | 93,292 | 35,814 | 269,745 |
| Data parsing and extraction | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 10.0 | 10.0 | 100.0% | 0 | 4.71s | 10,515 | 312 | 0 | |
| Trinity Large Thinking | 6.3 | 5.8 | 66.7% | 1 | 16.38s | 6,906 | 475 | 11,153 |
| Domain specific | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 4.1 | 4.4 | 44.5% | 2 | 19.30s | 990 | 61 | 8,940 | |
| Trinity Large Thinking | 5.3 | 7.2 | 44.4% | 1 | 124.17s | 756 | 99,334 | 172,018 |
| General Intelligence | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 10.0 | 10.0 | 100.0% | 0 | 5.62s | 714 | 325 | 0 | |
| Trinity Large Thinking | 10.0 | 10.0 | 100.0% | 0 | 32.22s | 501 | 7,035 | 7,774 |
| Instructions following | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 6.5 | 10.0 | 50.0% | 0 | 5.24s | 921 | 278 | 0 | |
| Trinity Large Thinking | 4.4 | 6.9 | 16.7% | 1 | 3.99s | 684 | 77 | 3,798 |
| Puzzle Solving | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 7.7 | 10.0 | 66.7% | 0 | 4.78s | 912 | 452 | 0 | |
| Trinity Large Thinking | 5.2 | 5.4 | 44.5% | 2 | 2.90s | 678 | 2,508 | 3,231 |
| Tool Calling | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 10.0 | 10.0 | 100.0% | 0 | 13.22s | 11,757 | 354 | 0 | |
| Trinity Large Thinking | 10.0 | 10.0 | 100.0% | 0 | 5.26s | 7,551 | 299 | 1,372 |
| Trivia | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 3.0 | 10.0 | 0.0% | 0 | 19.18s | 264 | 21 | 2,630 | |
| Trinity Large Thinking | 3.0 | 10.0 | 0.0% | 0 | 97.44s | 207 | 1,680 | 34,295 |
Quick Compare
Switch Comparison Pair
Trinity Large ThinkingmediumvsInklingnoneFree AvailableClaude Fable 5.1lowvsGemini 3.1 Flash Lite PreviewnoneTrinity Large ThinkingmediumvsQwen3.5 Plus 2026-04-20noneTrinity Large ThinkingmediumvsGemini 3.5 FlashnoneClaude Fable 5.1lowvsDeepSeek V4 Flash 0731noneTrinity Large ThinkingmediumvsKAT-Coder-Pro V2.5noneClaude Fable 5.1lowvsKAT-Coder-Pro V2.5mediumClaude Fable 5.1lowvsGemini 3.5 FlashnoneClaude Fable 5.1lowvsSpace Bunny AlphaxhighClaude Fable 5.1lowvsInklingnoneFree AvailableTrinity Large ThinkingmediumvsGemini 3.1 Flash Lite PreviewnoneClaude Fable 5.1lowvsQwen3.5 Plus 2026-04-20none