Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

MoonshotAI: Kimi K2.5 vs OpenAI: GPT-5 Nano

Last updated at: 2026-04-11

Metric Kimi K2.5 Kimi K2.5 none Release: 2026-01-27 GPT-5 Nano GPT-5 Nano medium Release: 2025-08-07
Score 5.5 6.3
Rank #72 #54
Consistency 8.7 6.5
Tests Correct
Attempt pass rate 40.7% 59.3%
Flaky tests 3 8
Total Runs 54 54
Cost per result 0.271 0.942
Total Cost $0.017 $0.066
Input Price $0.383 / 1M $0.050 / 1M
Output Price $1.720 / 1M $0.400 / 1M
Output Tokens 2,659 4,980
Reasoning Tokens 0 156,288
Response Time (avg) 13.37s 44.13s
Response Time (max) 42.13s 204.02s
Response Time (total) 147.05s 485.47s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 3.6 8.4 8.3% 1 6.24s 373 0
GPT-5 Nano 6.5 7.9 58.3% 1 25.50s 1,221 21,184
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 10.0 10.0 100.0% 0 38.78s 649 0
GPT-5 Nano 6.7 3.5 66.7% 1 40.73s 480 12,992
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 2.8 2.1 33.3% 1 19.16s 748 0
GPT-5 Nano 10.0 10.0 100.0% 0 65.96s 578 17,984
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 7.3 5.8 83.3% 1 42.13s 187 0
GPT-5 Nano 3.7 1.7 50.0% 2 21.42s 453 10,560
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 5.3 10.0 33.3% 0 4.38s 29 0
GPT-5 Nano 5.2 4.4 55.6% 2 204.02s 237 64,448
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 10.0 10.0 100.0% 0 4.00s 76 0
GPT-5 Nano 4.1 10.0 0.0% 0 17.51s 202 4,608
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 6.5 10.0 50.0% 0 2.67s 60 0
GPT-5 Nano 8.5 6.8 83.3% 1 11.90s 382 4,096
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 3.1 10.0 0.0% 0 4.73s 317 0
GPT-5 Nano 5.3 7.2 44.4% 1 19.81s 869 13,440
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.5 10.0 10.0 100.0% 0 13.99s 220 0
GPT-5 Nano 10.0 10.0 100.0% 0 33.30s 558 6,976

Quick Compare

Switch Comparison Pair