Navigasi
AI BENCHY
Advertise here

Kimi K2.7 Code (medium) vs GPT-5.3 Chat

Skor rata-rata hampir imbang di 7.5 vs 7.5. GPT-5.3 Chat memiliki biaya benchmark lebih rendah di $0.571 vs $0.692. GPT-5.3 Chat lebih cepat di 6.88s vs 84.25s, dengan tingkat keberhasilan 65.2% vs 68.2%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-07-25

Peringkat
#60
Total token output
262,507
Waktu respons (rata-rata)
84.25s
Total Biaya
$0.692
Peringkat
#62
Total token output
30,854
Waktu respons (rata-rata)
6.88s
Total Biaya
$0.571
Model yang direkomendasikan GPT-5.3 Chat

It has the best score here (7.5), while responding about 12.2x faster than Kimi K2.7 Code (medium).

Perbandingan terperinci

Metrik Kimi K2.7 Code Kimi K2.7 Code medium Rilis: 2026-06-12 GPT-5.3 Chat GPT-5.3 Chat none Rilis: 2026-03-03
Skor 7.5 7.5
Peringkat #60 #62
Keandalan 10.0 10.0
Konsistensi 8.3 8.2
Tes benar
Tingkat lulus per percobaan 65.2% 68.2%
Tes tidak stabil 4 5
Total Run 66 66
Biaya per hasil 6.457 4.387
Total Biaya $0.692 $0.571
Harga input $0.780 / 1M $1.750 / 1M
Harga output $3.500 / 1M $14.000 / 1M
Total token input 72,073 78,990
Token output 83,714 30,854
Token penalaran 178,793 0
Waktu respons (rata-rata) 84.25s 6.88s
Waktu respons (maks) 365.80s 18.33s
Waktu respons (total) 1769.22s 151.31s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#60 MoonshotAI: Kimi K2.7 Code

medium
Biaya
$0.025
Waktu
138.0s
Token
6,093 tok

#62 GPT-5.3 Chat

none
Biaya
$0.008
Waktu
8.1s
Token
634 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
Kimi K2.7 Code 7.8 9.3 66.7% 0 146.73s 4,650 1,864 25,635
GPT-5.3 Chat 5.6 4.7 55.6% 2 10.52s 7,302 6,632 0

Perbandingan Cepat

Ganti Pasangan Perbandingan