Navigasi
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Cobuddy vs Inception: Mercury 2

Cobuddy (medium) unggul dalam skor rata-rata dengan 4.7 vs 4.6. Cobuddy (medium) memiliki biaya benchmark lebih rendah di $0.000 vs $0.030. Mercury 2 lebih cepat di 829ms vs 39.90s, dengan tingkat keberhasilan 45.5% vs 22.7%.

Model yang direkomendasikanMercury 2Its score stays close to the best score here (4.6 vs 4.7), while responding about 48.1x faster than Cobuddy (medium).

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-07-18

Metrik Cobuddy Cobuddy medium Rilis: 2026-05-06 Mercury 2 Mercury 2 none Rilis: 2026-02-24
Skor 4.7 4.6
Peringkat #184 #185
Keandalan 10.0 10.0
Konsistensi 7.2 9.2
Tes benar
Tingkat lulus per percobaan 45.5% 22.7%
Tes tidak stabil 6 2
Total Run 63 66
Biaya per hasil 0.000 0.734
Total Biaya $0.000 $0.030
Harga input $0.000 / 1M $0.250 / 1M
Harga output $0.000 / 1M $0.750 / 1M
Total token input 37,449 88,704
Token output 1,677 9,564
Token penalaran 116,703 0
Waktu respons (rata-rata) 39.90s 829ms
Waktu respons (maks) 309.02s 4.52s
Waktu respons (total) 797.98s 18.24s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#184 Cobuddy

medium
No endpoints found for baidu/cobuddy:free.
Biaya
$0.000
Waktu
0.1s
Token
0 tok

#185 Mercury 2

none
Biaya
$0.002
Waktu
1.8s
Token
1,514 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Cobuddy 3.7 6.7 22.2% 1 79.17s 4,726 358 30,138
Mercury 2 3.4 9.6 0.0% 0 1.03s 7,229 3,088 0

Perbandingan Cepat

Ganti Pasangan Perbandingan