Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Compared models

Kimi K2.6 (medium) vs Kimi K2.5 (medium) vs GLM 5 (medium) vs Claude Opus 4.7 (medium) benchmark comparison: Claude Opus 4.7 (medium) leads on Score with 8.7. Kimi K2.6 (medium) leads on Reliability with 10.0. GLM 5 (medium) has the lowest Total Cost at $0.228. Claude Opus 4.7 (medium) is fastest at 7.62s.

Last updated at: 2026-09-10

Compared models

Rank
#128
Total Output Tokens
347,519
Response Time (avg)
115.89s
Total Cost
$1.247
Rank
#147
Total Output Tokens
220,942
Response Time (avg)
98.19s
Total Cost
$0.479
Rank
#92
Total Output Tokens
124,566
Response Time (avg)
33.54s
Total Cost
$0.228
Rank
#43
Total Output Tokens
29,990
Response Time (avg)
7.62s
Total Cost
$1.476
Recommended model Claude Opus 4.7 (medium)

It has the best score here (8.7), while responding about 10.8x faster than the other models in this comparison.

Detailed comparison

Metric Kimi K2.6 Kimi K2.6 medium Release: 2026-04-20 Kimi K2.5 Kimi K2.5 medium Release: 2026-01-27 GLM 5 GLM 5 medium Release: 2026-02-12 Claude Opus 4.7 Claude Opus 4.7 medium Release: 2026-04-16
Score 7.2 7.0 7.7 8.7
Rank #128 #147 #92 #43
Reliability 10.0 10.0 10.0 10.0
Consistency 8.3 7.0 8.1 9.6
Attempts 66/66 66/66 63/66 66/66
Tests Correct
Attempt pass rate 63.6% 63.6% 78.8% 83.3%
Flaky tests 4 8 4 1
Total Runs 66 66 63 66
Cost per result 10.029 4.849 1.668 8.200
Total Cost $1.247 $0.479 $0.228 $1.476
Input Price $0.950 / 1M $0.450 / 1M $0.600 / 1M $5.000 / 1M
Output Price $4.000 / 1M $2.250 / 1M $1.920 / 1M $25.000 / 1M
Total Input Tokens 68,913 118,496 35,224 145,249
Output Tokens 64,654 53,522 21,570 24,948
Reasoning Tokens 282,865 167,420 102,996 5,042
Response Time (avg) 115.89s 98.19s 33.54s 7.62s
Response Time (max) 876.20s 281.00s 99.85s 65.40s
Response Time (total) 2433.66s 1571.05s 435.99s 159.94s
Parameters 1T total (32B active) 1T total (32B active) 744B total (40B active) ~5T total (~500B active)
Availability Weights available Weights available Open source Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#128 MoonshotAI: Kimi K2.6

medium
Cost
$0.013
Time
103.4s
Tokens
3,620 tok

#147 MoonshotAI: Kimi K2.5

medium
Cost
$0.030
Time
58.6s
Tokens
8,683 tok

#92 GLM 5

medium
Cost
$0.005
Time
20.7s
Tokens
2,068 tok

#43 Claude Opus 4.7

medium
Cost
$0.059
Time
26.8s
Tokens
2,475 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Kimi K2.6 5.7 8.6 33.3% 0 214.42s 2,925 9,970 77,189
Kimi K2.5 6.1 4.6 66.7% 2 217.49s 6,935 5,705 74,693
GLM 5 10.0 10.0 100.0% 0 74.30s 7,254 2,997 52,930
Claude Opus 4.7 7.6 7.2 77.8% 1 12.96s 10,635 7,629 1,114

Quick Compare

Switch Comparison Pair