Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

MiniMax: MiniMax M3 vs OpenAI: GPT-5.3 Chat

Last updated at: 2026-06-01

Metric MiniMax M3 MiniMax M3 medium Release: 2026-06-01 GPT-5.3 Chat GPT-5.3 Chat none Release: 2026-03-03
Score 7.3 7.4
Rank #65 #57
Reliability 9.6 10.0
Consistency 8.4 8.4
Tests Correct
Attempt pass rate 68.3% 68.3%
Flaky tests 6 4
Total Runs 60 60
Cost per result 1.083 3.350
Total Cost $0.120 $0.402
Input Price $0.300 / 1M $1.750 / 1M
Output Price $1.200 / 1M $14.000 / 1M
Output Tokens 46,884 24,757
Reasoning Tokens 85,935 0
Response Time (avg) 68.44s 6.13s
Response Time (max) 431.03s 18.33s
Response Time (total) 1300.32s 122.61s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 5.5 3.7 66.7% 3 14.95s 874 3,414
GPT-5.3 Chat 6.7 8.1 58.3% 1 3.86s 3,167 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 7.5 10.0 66.7% 1 185.58s 4,071 26,059
GPT-5.3 Chat 6.9 6.2 66.7% 1 10.52s 4,772 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 10.0 10.0 100.0% 0 65.30s 1,306 6,253
GPT-5.3 Chat 10.0 10.0 100.0% 0 11.96s 2,614 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 10.0 10.0 100.0% 0 14.92s 514 3,164
GPT-5.3 Chat 10.0 10.0 100.0% 0 2.21s 942 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 6.0 10.0 44.4% 1 233.13s 16,254 19,070
GPT-5.3 Chat 3.5 4.4 33.3% 2 13.01s 8,264 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 5.1 3.4 33.3% 1 33.25s 2,487 2,523
GPT-5.3 Chat 4.6 10.0 0.0% 0 1.99s 319 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 9.8 10.0 100.0% 0 6.14s 103 920
GPT-5.3 Chat 9.8 10.0 100.0% 0 3.51s 1,491 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 7.9 9.9 66.7% 0 49.91s 11,946 13,761
GPT-5.3 Chat 10.0 10.0 100.0% 0 2.99s 1,758 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 10.0 10.0 100.0% 0 11.91s 281 555
GPT-5.3 Chat 10.0 10.0 100.0% 0 8.36s 861 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
MiniMax M3 3.0 10.0 0.0% 0 100.80s 9,048 10,216
GPT-5.3 Chat 3.0 10.0 0.0% 0 4.38s 569 0

Quick Compare

Switch Comparison Pair