Navigate
AI BENCHY
Advertise here

Kwaipilot: KAT-Coder-Pro V2.5 vs Meituan: LongCat 2.0

The average score is effectively tied at 7.4 vs 7.4. KAT-Coder-Pro V2.5 (low) has the lower benchmark cost at $0.387 vs $0.478. KAT-Coder-Pro V2.5 (low) is faster at 19.47s vs 136.64s, with pass rates of 69.7% vs 60.6%.

Recommended modelKAT-Coder-Pro V2.5 (low)It has the best score here (7.4), while responding about 7.0x faster than LongCat 2.0 (medium).

Last updated at: 2026-07-20

Metric KAT-Coder-Pro V2.5 KAT-Coder-Pro V2.5 low Release: 2026-07-14 LongCat 2.0 LongCat 2.0 medium Release: 2026-07-20
Score 7.4 7.4
Rank #62 #60
Reliability 10.0 9.8
Consistency 7.0 8.9
Tests Correct
Attempt pass rate 69.7% 60.6%
Flaky tests 8 3
Total Runs 66 66
Cost per result 3.514 3.978
Total Cost $0.387 $0.478
Input Price $0.740 / 1M $0.300 / 1M
Output Price $2.960 / 1M $1.200 / 1M
Total Input Tokens 87,673 97,315
Output Tokens 7,166 44,384
Reasoning Tokens 101,474 329,012
Response Time (avg) 19.47s 136.64s
Response Time (max) 209.15s 939.52s
Response Time (total) 428.31s 3006.01s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#62 KAT-Coder-Pro V2.5

low
Cost
$0.016
Time
47.6s
Tokens
5,182 tok

#60 LongCat 2.0

medium
Cost
$0.016
Time
285.1s
Tokens
13,057 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
KAT-Coder-Pro V2.5 7.8 10.0 66.7% 0 24.87s 7,893 402 24,945
LongCat 2.0 10.0 10.0 100.0% 0 455.00s 7,419 533 214,143

Quick Compare

Switch Comparison Pair