Navigasi
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Trinity Large Thinking (high) vs Ring 2.6 1t (medium)

Skor rata-rata hampir imbang di 5.7 vs 5.7. Ring 2.6 1t (medium) memiliki biaya benchmark lebih rendah di $0.072 vs $0.645. Ring 2.6 1t (medium) lebih cepat di 65.52s vs 74.52s, dengan tingkat keberhasilan 42.0% vs 56.5%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-10-01

Model yang Dibandingkan

Peringkat
#253
Total token output
943,044
Waktu respons (rata-rata)
74.52s
Total Biaya
$0.645
Peringkat
#252
Total token output
165,031
Waktu respons (rata-rata)
65.52s
Total Biaya
$0.072
Model yang direkomendasikan Ring 2.6 1t (medium)

It has the best score here (5.7), while costing about 9.0x less than Trinity Large Thinking (high).

Perbandingan terperinci

Metrik Trinity Large Thinking Trinity Large Thinking high Rilis: 2026-07-28 Ring 2.6 1t Ring 2.6 1t medium Rilis: 2026-05-10
Skor 5.7 5.7
Peringkat #253 #252
Keandalan 10.0 10.0
Konsistensi 7.3 8.9
Percobaan 69/69 69/69
Tes benar
Tingkat lulus per percobaan 42.0% 56.5%
Tes tidak stabil 8 3
Total Run 69 69
Biaya per hasil 11.289 0.654
Total Biaya $0.645 $0.072
Harga input $0.250 / 1M $0.000 / 1M
Harga output $0.800 / 1M $0.000 / 1M
Total token input 283,116 113,613
Token output 277,920 122,706
Token penalaran 665,124 42,325
Waktu respons (rata-rata) 74.52s 65.52s
Waktu respons (maks) 510.21s 304.19s
Waktu respons (total) 1713.90s 1376.01s
Parameter 398B total (13B aktif) 1T total (63B aktif)
Ketersediaan Bobot tersedia Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#253 Trinity Large Thinking

high
Biaya
$0.028
Waktu
130.1s
Token
34,387 tok

#252 Ring 2.6 1t

medium
Ring-2.6-1T is no longer available as a free model. It has transitioned to a paid model. Continue using it here: https://openrouter.ai/inclusionai/ring-2.6-1t
Biaya
$0.000
Waktu
0.1s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Trinity Large Thinking 3.7 4.7 33.3% 2 245.04s 7,204 83,616 266,836
Ring 2.6 1t 5.3 10.0 33.3% 0 59.65s 834 1,369 3,985

Perbandingan Cepat

Ganti Pasangan Perbandingan