Navigate
Advertise here

KAT-Coder-Pro V2.5 (high) vs GLM 5.2

The average score is effectively tied at 6.5 vs 6.5. GLM 5.2 has the lower benchmark cost at $0.161 vs $0.506. GLM 5.2 is faster at 11.51s vs 21.21s, with pass rates of 60.9% vs 62.3%.

Last updated at: 2026-10-01

Compared models

Rank
#198
Total Output Tokens
144,359
Response Time (avg)
21.21s
Total Cost
$0.506
Rank
#197
Total Output Tokens
19,230
Response Time (avg)
11.51s
Total Cost
$0.161
Recommended model GLM 5.2

It has the best score here (6.5), while costing about 3.2x less than KAT-Coder-Pro V2.5 (high).

Detailed comparison

Metric KAT-Coder-Pro V2.5 KAT-Coder-Pro V2.5 high Release: 2026-07-14 GLM 5.2 GLM 5.2 none Release: 2026-06-17
Score 6.5 6.5
Rank #198 #197
Reliability 10.0 10.0
Consistency 7.9 9.0
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 60.9% 62.3%
Flaky tests 6 3
Total Runs 69 69
Cost per result 4.599 1.825
Total Cost $0.506 $0.161
Input Price $0.740 / 1M $0.325 / 1M
Output Price $2.960 / 1M $3.990 / 1M
Total Input Tokens 106,085 257,698
Output Tokens 9,071 19,230
Reasoning Tokens 135,288 0
Response Time (avg) 21.21s 11.51s
Response Time (max) 199.97s 79.65s
Response Time (total) 487.83s 264.72s
Parameters ~700B total (~72B active) 744B total (40B active)
Availability Closed Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#198 KAT-Coder-Pro V2.5

high
Cost
$0.009
Time
26.4s
Tokens
3,117 tok

#197 GLM 5.2

none
Cost
$0.033
Time
87.7s
Tokens
7,455 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
KAT-Coder-Pro V2.5 6.4 7.9 44.4% 1 22.00s 7,893 422 20,461
GLM 5.2 3.7 9.5 0.0% 0 7.55s 7,263 1,958 0

Quick Compare

Switch Comparison Pair