Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Trinity Large Preview vs Qwen3 Coder Next (medium)

Trinity Large Preview unggul dalam skor rata-rata dengan 4.8 vs 4.6. Trinity Large Preview memiliki biaya benchmark lebih rendah di $0.008 vs $0.034. Trinity Large Preview lebih cepat di 2.98s vs 9.07s, dengan tingkat keberhasilan 21.2% vs 25.8%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-09-02

Peringkat
#273
Total token output
2,169
Waktu respons (rata-rata)
2.98s
Total Biaya
$0.008
Peringkat
#280
Total token output
19,069
Waktu respons (rata-rata)
9.07s
Total Biaya
$0.034
Model yang direkomendasikan Trinity Large Preview

It has the best score here (4.8), while costing about 4.3x less than Qwen3 Coder Next (medium).

Perbandingan terperinci

Metrik Trinity Large Preview Trinity Large Preview none Rilis: 2026-01-27 Qwen3 Coder Next Qwen3 Coder Next medium Rilis: 2026-02-03
Skor 4.8 4.6
Peringkat #273 #280
Keandalan 10.0 10.0
Konsistensi 8.9 8.6
Percobaan 63/66 66/66
Tes benar
Tingkat lulus per percobaan 21.2% 25.8%
Tes tidak stabil 2 4
Total Run 63 66
Biaya per hasil 0.017 1.058
Total Biaya $0.008 $0.034
Harga input $0.243 / 1M $0.120 / 1M
Harga output $0.243 / 1M $0.800 / 1M
Total token input 29,828 148,203
Token output 2,169 19,069
Token penalaran 0 0
Waktu respons (rata-rata) 2.98s 9.07s
Waktu respons (maks) 14.34s 81.80s
Waktu respons (total) 56.57s 154.19s
Parameter 398B total (13B aktif) 80B total (3B aktif)
Ketersediaan Sumber terbuka Sumber terbuka

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#273 Trinity Large Preview

none
No endpoints found for arcee-ai/trinity-large-preview:free.
Biaya
$0.000
Waktu
0.0s
Token
0 tok

#280 Qwen3 Coder Next

medium
SVG tidak valid
Biaya
$0.000
Waktu
300.0s
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Trinity Large Preview 3.7 7.7 11.1% 1 14.34s 738 397 0
Qwen3 Coder Next 3.7 7.2 22.2% 1 924ms 7,185 336 0

Perbandingan Cepat

Ganti Pasangan Perbandingan