Navigate
Advertise here

Trinity Large Preview vs GLM 4.7 Flash (medium)

Trinity Large Preview leads on average score with 4.8 vs 4.3. Trinity Large Preview has the lower benchmark cost at $0.008 vs $0.168. Trinity Large Preview is faster at 2.98s vs 134.34s, with pass rates of 21.2% vs 31.8%.

Last updated at: 2026-09-22

Compared models

Rank
#306
Total Output Tokens
2,169
Response Time (avg)
2.98s
Total Cost
$0.008
Rank
#319
Total Output Tokens
419,534
Response Time (avg)
134.34s
Total Cost
$0.168
Recommended model Trinity Large Preview

It has the best score here (4.8), while costing about 21.6x less than GLM 4.7 Flash (medium).

Detailed comparison

Metric Trinity Large Preview Trinity Large Preview none Release: 2026-01-27 GLM 4.7 Flash GLM 4.7 Flash medium Release: 2026-01-19
Score 4.8 4.3
Rank #306 #319
Reliability 10.0 8.6
Consistency 8.9 7.0
Attempts 63/66 66/66
Tests Correct
Attempt pass rate 21.2% 31.8%
Flaky tests 2 8
Total Runs 63 66
Cost per result 0.017 4.180
Total Cost $0.008 $0.168
Input Price $0.243 / 1M $0.061 / 1M
Output Price $0.243 / 1M $0.400 / 1M
Total Input Tokens 29,828 79,104
Output Tokens 2,169 42,254
Reasoning Tokens 0 377,280
Response Time (avg) 2.98s 134.34s
Response Time (max) 14.34s 1539.97s
Response Time (total) 56.57s 2015.13s
Parameters 398B total (13B active) 30B total (3B active)
Availability Open source Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#306 Trinity Large Preview

none
No endpoints found for arcee-ai/trinity-large-preview:free.
Cost
$0.000
Time
0.0s
Tokens
0 tok

#319 GLM 4.7 Flash

medium
No output was saved. The original provider response or failure reason is unavailable.
Cost
$0.000
Time
186.2s
Tokens
12,112 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Trinity Large Preview 3.7 7.7 11.1% 1 14.34s 738 397 0
GLM 4.7 Flash 3.2 7.4 11.1% 1 55.33s 3,106 4,981 22,387

Quick Compare

Switch Comparison Pair