ナビゲーション
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Trinity Large Thinking (medium) vs Inkling

平均スコアは 6.1 vs 6.1 でほぼ同等です。 Inkling の benchmark コストが低く、$0.304 vs $0.820 です。 Inkling の方が高速で、5.30s vs 84.01s です、成功率は 49.3% vs 31.9% です。

ベンチマークは AI BENCHY テストスイートから次の日時に生成: 2026-10-01

比較対象モデル

順位
#232
合計出力トークン
1,033,971
応答時間(平均)
84.01s
合計コスト
$0.820
順位
#233
合計出力トークン
16,536
応答時間(平均)
5.30s
合計コスト
$0.304
おすすめモデル Inkling

ここでは最高スコア(6.1)で、Trinity Large Thinking (medium) より約 2.7 倍低コストです。

詳細比較

指標 Trinity Large Thinking Trinity Large Thinking medium リリース: 2026-07-28 Inkling Inkling none リリース: 2026-07-18 無料で利用可能
スコア 6.1 6.1
順位 #232 #233
信頼性 10.0 10.0
一貫性 7.7 9.6
試行回数 69/69 69/69
正解テスト
試行ごとの合格率 49.3% 31.9%
不安定なテスト 7 1
総実行回数 69 69
結果あたりのコスト 10.772 4.337
合計コスト $0.820 $0.304
入力価格 $0.250 / 1M $1.000 / 1M
出力価格 $0.800 / 1M $4.050 / 1M
合計入力トークン 319,352 236,599
出力トークン 158,447 16,536
推論トークン 875,524 0
応答時間(平均) 84.01s 5.30s
応答時間(最大) 525.14s 48.02s
応答時間(合計) 1932.24s 121.91s
パラメータ 398B 総数 (13B アクティブ) 975B 総数 (41B アクティブ)
公開状況 重み公開 オープンソース

モデル生成ショーケース

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#232 Trinity Large Thinking

medium
コスト
$0.016
時間
75.9s
トークン
19,401 tok

#233 Thinking Machines: Inkling

none
Provider returned error
コスト
$0.000
時間
5.3s
トークン
0 tok

スコア上位モデル

スコア vs 総コスト

応答時間(平均)

スコア vs 応答時間(平均)

合計出力トークン

スコア vs 合計出力トークン

カテゴリ内訳

コーディング スコア 一貫性 試行ごとの合格率 不安定なテスト 正解テスト 応答時間(平均) 入力トークン 出力トークン 推論トークン
Trinity Large Thinking 7.5 10.0 66.7% 0 179.02s 7,248 7,413 351,155
Inkling 4.5 10.0 0.0% 0 1.01s 7,356 436 0

クイック比較

比較ペアを切り替え