Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

MiniMax: MiniMax M2.7 vs OpenAI: GPT-4o-mini

Last updated at: 2026-05-10

Metric MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18 GPT-4o-mini GPT-4o-mini none Release: 2024-07-18
Score 5.1 4.9
Rank #125 #129
Reliability 10.0 10.0
Consistency 5.7 9.9
Tests Correct
Attempt pass rate 49.1% 26.3%
Flaky tests 10 0
Total Runs 57 57
Cost per result 2.367 0.099
Total Cost $0.095 $0.005
Input Price $0.300 / 1M $0.150 / 1M
Output Price $1.200 / 1M $0.600 / 1M
Output Tokens 8,052 1,962
Reasoning Tokens 66,239 0
Response Time (avg) 30.62s 1.90s
Response Time (max) 117.04s 7.58s
Response Time (total) 551.14s 22.79s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 7.9 6.3 83.3% 2 40.32s 3,010 17,716
GPT-4o-mini 4.8 10.0 25.0% 0 1.34s 186 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 10.0 10.0 100.0% 0 91.27s 467 15,175
GPT-4o-mini 3.0 8.7 0.0% 0 2.55s 347 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 41.03s 369 4,480
GPT-4o-mini 3.0 10.0 0.0% 0 7.58s 568 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 6.3 5.8 66.7% 1 21.95s 187 5,882
GPT-4o-mini 10.0 10.0 100.0% 0 1.27s 183 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 19.00s 8 2,796
GPT-4o-mini 3.0 10.0 0.0% 0 637ms 15 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.9 2.5 33.3% 1 38.70s 92 5,204
GPT-4o-mini 4.0 10.0 0.0% 0 909ms 66 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.8 1.6 50.0% 2 12.64s 213 2,457
GPT-4o-mini 6.3 10.0 50.0% 0 1.27s 69 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.5 4.4 33.3% 2 25.62s 334 8,076
GPT-4o-mini 3.5 10.0 0.0% 0 1.30s 308 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 12.05s 304 1,001
GPT-4o-mini 10.0 10.0 100.0% 0 2.51s 205 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 22.77s 3,068 3,452
GPT-4o-mini 3.0 10.0 0.0% 0 794ms 15 0

Quick Compare

Switch Comparison Pair