Navigasi
AI BENCHY
Advertise here

Kimi K2.7 Code (medium) vs Grok 4.6 (low)

Kimi K2.7 Code (medium) unggul dalam skor rata-rata dengan 7.4 vs 7.4. Grok 4.6 (low) memiliki biaya benchmark lebih rendah di $0.449 vs $0.709. Grok 4.6 (low) lebih cepat di 14.55s vs 80.09s, dengan tingkat keberhasilan 63.6% vs 69.7%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-08-14

Peringkat
#89
Total token output
239,546
Waktu respons (rata-rata)
80.09s
Total Biaya
$0.709
Peringkat
#93
Total token output
38,048
Waktu respons (rata-rata)
14.55s
Total Biaya
$0.449
Model yang direkomendasikan Grok 4.6 (low)

Its score stays close to the best score here (7.4 vs 7.4), while costing about 1.6x less than Kimi K2.7 Code (medium).

Perbandingan terperinci

Metrik Kimi K2.7 Code Kimi K2.7 Code medium Rilis: 2026-06-12 Grok 4.6 Grok 4.6 low Rilis: 2026-08-12
Skor 7.4 7.4
Peringkat #89 #93
Keandalan 10.0 10.0
Konsistensi 8.0 9.6
Percobaan 66/66 66/66
Tes benar
Tingkat lulus per percobaan 63.6% 69.7%
Tes tidak stabil 5 1
Total Run 66 66
Biaya per hasil 7.096 2.989
Total Biaya $0.709 $0.449
Harga input $0.710 / 1M $2.000 / 1M
Harga output $3.500 / 1M $6.000 / 1M
Total token input 72,085 109,960
Token output 57,453 5,124
Token penalaran 182,093 32,924
Waktu respons (rata-rata) 80.09s 14.55s
Waktu respons (maks) 365.80s 26.46s
Waktu respons (total) 1681.85s 320.17s
Parameter 1T total (32B aktif) ~1.7T total (~170B aktif)
Ketersediaan Bobot tersedia Tertutup

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#89 MoonshotAI: Kimi K2.7 Code

medium
Biaya
$0.025
Waktu
138.0s
Token
6,093 tok

#93 SpaceXAI: Grok 4.6

low
Biaya
$0.011
Waktu
26.0s
Token
1,941 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Kimi K2.7 Code 7.8 9.3 66.7% 0 146.73s 4,650 1,864 25,635
Grok 4.6 5.5 7.1 44.4% 1 21.69s 9,579 349 7,740

Perbandingan Cepat

Ganti Pasangan Perbandingan