Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Google: Gemini 3.1 Flash Lite Preview vs MiniMax: MiniMax M2.5

Summary

Gemini 3.1 Flash Lite Preview vs MiniMax M2.5 benchmark comparison: Gemini 3.1 Flash Lite Preview leads on average score with 6.4 vs 4.7. Gemini 3.1 Flash Lite Preview has the lower benchmark cost at $0.018 vs $0.303. Gemini 3.1 Flash Lite Preview is faster at 1.21s vs 65.37s, with pass rates of 60.3% vs 46.0%.

Recommended model: Gemini 3.1 Flash Lite Preview - It has the best score here (6.4), while costing about 17.1x less than MiniMax M2.5.

Last updated at: 2026-06-12

Metric Gemini 3.1 Flash Lite Preview Gemini 3.1 Flash Lite Preview none Release: 2026-03-03 MiniMax M2.5 MiniMax M2.5 medium Release: 2026-02-12
Score 6.4 4.7
Rank #82 #151
Reliability 10.0 10.0
Consistency 9.7 6.5
Tests Correct
Attempt pass rate 60.3% 46.0%
Flaky tests 1 9
Total Runs 63 63
Cost per result 0.148 7.900
Total Cost $0.018 $0.303
Input Price $0.250 / 1M $0.150 / 1M
Output Price $1.500 / 1M $0.900 / 1M
Total Input Tokens 37,582 43,706
Output Tokens 5,547 109,495
Reasoning Tokens 0 330,814
Response Time (avg) 1.21s 65.37s
Response Time (max) 3.39s 251.36s
Response Time (total) 25.45s 849.76s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#82 Gemini 3.1 Flash Lite Preview

none
Cost
$0.003
Time
4.7s
Tokens
1,827 tok

#151 MiniMax M2.5

medium
Invalid SVG
Cost
$0.000
Time
300.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 7.5 8.4 66.7% 1 1.04s 504 1,092 0
MiniMax M2.5 7.9 6.3 83.3% 2 20.82s 612 286 45,344
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 5.5 10.0 33.3% 0 967ms 8,128 670 0
MiniMax M2.5 3.4 9.1 0.0% 0 188.58s 6,076 357 106,177
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 3.20s 13,026 339 0
MiniMax M2.5 4.5 2.1 66.7% 1 60.39s 21,104 740 9,713
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.22s 7,550 399 0
MiniMax M2.5 4.6 1.7 66.7% 2 7.48s 6,584 266 3,835
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 5.3 10.0 33.3% 0 942ms 641 568 0
MiniMax M2.5 2.9 4.4 22.2% 2 237.27s 308 105,047 133,487
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 4.0 10.0 0.0% 0 741ms 488 69 0
MiniMax M2.5 3.8 2.5 33.3% 1 6.63s 492 25 1,686
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.13s 623 574 0
MiniMax M2.5 7.5 10.0 50.0% 0 621ms 699 156 1,495
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 900ms 570 1,045 0
MiniMax M2.5 5.3 7.2 44.4% 1 11.21s 495 1,069 9,605
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 3.39s 5,894 782 0
MiniMax M2.5 10.0 10.0 100.0% 0 15.35s 7,123 269 937
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 814ms 158 9 0
MiniMax M2.5 3.0 10.0 0.0% 0 80.79s 213 1,280 18,535

Quick Compare

Switch Comparison Pair