Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

MiniMax: MiniMax M2.7 vs MoonshotAI: Kimi K2.5

Last updated at: 2026-03-18

Metric MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18 Kimi K2.5 Kimi K2.5 none Release: 2026-01-27
Score 5.0 5.3
Rank #64 #59
Consistency 5.3 8.7
Tests Correct
Attempt pass rate 49.0% 37.3%
Flaky tests 10 3
Total Runs 51 51
Cost per result 2.398 0.297
Total Cost $0.072 $0.015
Input Price $0.300 / 1M $0.450 / 1M
Output Price $1.200 / 1M $2.200 / 1M
Output Tokens 4,517 2,010
Reasoning Tokens 47,612 0
Response Time (avg) 27.32s 10.83s
Response Time (max) 117.04s 42.13s
Response Time (total) 437.10s 108.27s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 7.9 6.3 83.3% 2 40.32s 3,010 17,716
Kimi K2.5 3.6 8.4 8.3% 1 6.24s 373 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 41.03s 369 4,480
Kimi K2.5 2.8 2.1 33.3% 1 19.16s 748 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 6.3 5.8 66.7% 1 21.95s 187 5,882
Kimi K2.5 7.3 5.8 83.3% 1 42.13s 187 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 19.00s 8 2,796
Kimi K2.5 5.3 10.0 33.3% 0 4.38s 29 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.9 2.5 33.3% 1 38.70s 92 5,204
Kimi K2.5 10.0 10.0 100.0% 0 4.00s 76 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.7 1.8 50.0% 2 12.64s 213 2,457
Kimi K2.5 6.5 10.0 50.0% 0 2.67s 60 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.8 4.5 33.3% 2 25.62s 334 8,076
Kimi K2.5 3.1 10.0 0.0% 0 4.73s 317 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 12.05s 304 1,001
Kimi K2.5 10.0 10.0 100.0% 0 13.99s 220 0

Quick Compare

Switch Comparison Pair