Navigate
Advertise here

KAT-Coder-Air V2.5 vs GPT-5.4 Nano

The average score is effectively tied at 4.8 vs 4.8. GPT-5.4 Nano has the lower benchmark cost at $0.041 vs $0.070. GPT-5.4 Nano is faster at 2.56s vs 12.46s, with pass rates of 36.4% vs 27.3%.

Last updated at: 2026-09-10

Compared models

Rank
#290
Total Output Tokens
97,752
Response Time (avg)
12.46s
Total Cost
$0.070
Rank
#287
Total Output Tokens
13,794
Response Time (avg)
2.56s
Total Cost
$0.041
Recommended model GPT-5.4 Nano

It has the best score here (4.8), while costing about 1.7x less than KAT-Coder-Air V2.5.

Detailed comparison

Metric KAT-Coder-Air V2.5 KAT-Coder-Air V2.5 none Release: 2026-07-14 GPT-5.4 Nano GPT-5.4 Nano none Release: 2026-03-17
Score 4.8 4.8
Rank #290 #287
Reliability 10.0 10.0
Consistency 7.2 8.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 36.4% 27.3%
Flaky tests 7 4
Total Runs 66 66
Cost per result 1.382 1.011
Total Cost $0.070 $0.041
Input Price $0.150 / 1M $0.200 / 1M
Output Price $0.600 / 1M $1.250 / 1M
Total Input Tokens 69,376 115,933
Output Tokens 97,752 13,794
Reasoning Tokens 0 0
Response Time (avg) 12.46s 2.56s
Response Time (max) 121.05s 25.50s
Response Time (total) 274.11s 56.37s
Parameters ~35B total (~3B active) ~120B total (~5B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#290 KAT-Coder-Air V2.5

none
Cost
$0.002
Time
14.5s
Tokens
2,619 tok

#287 GPT-5.4 Nano

none
Cost
$0.008
Time
46.1s
Tokens
5,735 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
KAT-Coder-Air V2.5 3.3 9.6 0.0% 0 14.31s 6,472 16,918 0
GPT-5.4 Nano 4.6 7.9 22.2% 1 2.22s 7,305 613 0

Quick Compare

Switch Comparison Pair