নেভিগেশন
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Trinity Large Preview vs Qwen: Qwen3.5-9B

Trinity Large Preview average score-এ এগিয়ে: 4.8 vs 3.8. Trinity Large Preview-এর benchmark খরচ কম: $0.008 vs $0.036. Trinity Large Preview দ্রুত: 2.98s vs 82.24s, pass rates 21.2% vs 25.8%.

প্রস্তাবিত মডেলTrinity Large PreviewIt has the best score here (4.8), while costing about 4.6x less than Qwen3.5-9B (medium).

AI BENCHY টেস্ট স্যুট থেকে বেঞ্চমার্ক তৈরি হয়েছে: 2026-07-22

মেট্রিক Trinity Large Preview Trinity Large Preview none প্রকাশ: 2026-01-27 Qwen3.5-9B Qwen3.5-9B medium প্রকাশ: 2026-03-02
স্কোর 4.8 3.8
র‍্যাঙ্ক #192 #214
নির্ভরযোগ্যতা 10.0 5.0
ধারাবাহিকতা 8.9 8.1
সঠিক টেস্ট
প্রতি চেষ্টায় পাস রেট 21.2% 25.8%
অস্থির টেস্ট 2 5
মোট রান 63 66
প্রতি ফলাফলে খরচ 0.017 1.187
মোট খরচ $0.008 $0.036
ইনপুট মূল্য $0.243 / 1M $0.100 / 1M
আউটপুট মূল্য $0.243 / 1M $0.150 / 1M
মোট ইনপুট টোকেন 29,828 17,070
আউটপুট টোকেন 2,169 29,045
রিজনিং টোকেন 0 209,516
প্রতিক্রিয়া সময় (গড়) 2.98s 82.24s
প্রতিক্রিয়া সময় (সর্বোচ্চ) 14.34s 226.38s
প্রতিক্রিয়া সময় (মোট) 56.57s 1315.88s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#192 Trinity Large Preview

none
No endpoints found for arcee-ai/trinity-large-preview:free.
খরচ
$0.000
সময়
0.0s
টোকেন
0 tok

#214 Qwen3.5-9B

medium
খরচ
$0.001
সময়
35.9s
টোকেন
3,030 tok

স্কোর অনুযায়ী শীর্ষ মডেল

স্কোর বনাম মোট খরচ

প্রতিক্রিয়া সময় (গড়)

স্কোর vs প্রতিক্রিয়া সময় (গড়)

মোট আউটপুট টোকেন

স্কোর vs মোট আউটপুট টোকেন

বিভাগভিত্তিক বিশ্লেষণ

কোডিং স্কোর ধারাবাহিকতা প্রতি চেষ্টায় পাস রেট অস্থির টেস্ট সঠিক টেস্ট প্রতিক্রিয়া সময় (গড়) ইনপুট টোকেন আউটপুট টোকেন রিজনিং টোকেন
Trinity Large Preview 3.7 7.7 11.1% 1 14.34s 738 397 0
Qwen3.5-9B 2.9 10.0 0.0% 0 100.88s 2,396 7,890 41,129

দ্রুত তুলনা

তুলনার জুটি বদলান