Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

MiniMax: MiniMax M2.7 vs MiniMax: MiniMax M3

Last updated at: 2026-06-02

Metric MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18 MiniMax M3 MiniMax M3 medium Release: 2026-06-01
Score 5.4 7.3
Rank #129 #65
Reliability 10.0 9.6
Consistency 6.7 8.4
Tests Correct
Attempt pass rate 48.3% 68.3%
Flaky tests 8 6
Total Runs 60 60
Cost per result 2.076 1.083
Total Cost $0.103 $0.120
Input Price $0.260 / 1M $0.300 / 1M
Output Price $1.200 / 1M $1.200 / 1M
Total Input Tokens 33,493 43,447
Output Tokens 8,224 46,884
Reasoning Tokens 73,373 85,935
Response Time (avg) 29.86s 68.44s
Response Time (max) 117.04s 431.03s
Response Time (total) 567.39s 1300.32s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 7.9 6.3 83.3% 2 40.32s 654 3,010 17,716
MiniMax M3 5.5 3.7 66.7% 3 14.95s 2,526 874 3,414
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 6.7 9.6 50.0% 0 54.73s 2,083 474 22,402
MiniMax M3 7.5 10.0 66.7% 1 185.58s 2,705 4,071 26,059
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 41.03s 14,233 369 4,480
MiniMax M3 10.0 10.0 100.0% 0 65.30s 14,760 1,306 6,253
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 6.3 5.8 66.7% 1 21.95s 7,152 187 5,882
MiniMax M3 10.0 10.0 100.0% 0 14.92s 8,088 514 3,164
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 19.00s 245 8 2,796
MiniMax M3 6.0 10.0 44.4% 1 233.13s 869 16,254 19,070
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.9 2.5 33.3% 1 38.70s 486 92 5,204
MiniMax M3 5.1 3.4 33.3% 1 33.25s 954 2,487 2,523
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.8 5.8 33.3% 1 12.80s 687 350 2,600
MiniMax M3 9.8 10.0 100.0% 0 6.14s 1,623 103 920
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 5.9 7.2 55.6% 1 24.87s 675 362 7,840
MiniMax M3 7.9 9.9 66.7% 0 49.91s 2,079 11,946 13,761
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 12.05s 7,067 304 1,001
MiniMax M3 10.0 10.0 100.0% 0 11.91s 9,168 281 555
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 22.77s 211 3,068 3,452
MiniMax M3 3.0 10.0 0.0% 0 100.80s 675 9,048 10,216

Quick Compare

Switch Comparison Pair