Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

MiniMax: MiniMax M2.7 vs Qwen: Qwen3.6 35B A3B

Summary

MiniMax M2.7 vs Qwen3.6 35B A3B benchmark comparison: MiniMax M2.7 leads on average score with 5.3 vs 4.6. Qwen3.6 35B A3B has the lower benchmark cost at $0.031 vs $0.113. Qwen3.6 35B A3B is faster at 3.73s vs 38.18s, with pass rates of 46.0% vs 30.2%.

Recommended model: Qwen3.6 35B A3B - Its score stays close to the best score here (4.6 vs 5.3), while costing about 3.7x less than MiniMax M2.7.

Last updated at: 2026-06-10

Metric MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18 Qwen3.6 35B A3B Qwen3.6 35B A3B none Release: 2026-04-20
Score 5.3 4.6
Rank #131 #154
Reliability 10.0 10.0
Consistency 6.8 8.0
Tests Correct
Attempt pass rate 46.0% 30.2%
Flaky tests 8 5
Total Runs 63 63
Cost per result 2.494 0.754
Total Cost $0.113 $0.031
Input Price $0.270 / 1M $0.140 / 1M
Output Price $1.080 / 1M $1.000 / 1M
Total Input Tokens 34,371 19,329
Output Tokens 8,981 27,755
Reasoning Tokens 89,812 0
Response Time (avg) 38.18s 3.73s
Response Time (max) 196.21s 22.52s
Response Time (total) 763.60s 70.86s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#131 MiniMax M2.7

medium
Cost
$0.022
Time
22.8s
Tokens
9,250 tok

#154 Qwen3.6 35B A3B

none
Cost
$0.008
Time
30.1s
Tokens
6,317 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 7.9 6.3 83.3% 2 40.32s 654 3,010 17,716
Qwen3.6 35B A3B 3.6 7.6 16.7% 1 2.10s 696 1,571 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 5.7 9.1 33.3% 0 101.89s 2,961 1,231 38,841
Qwen3.6 35B A3B 5.5 10.0 33.3% 0 8.77s 7,911 11,161 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 41.03s 14,233 369 4,480
Qwen3.6 35B A3B 3.0 10.0 0.0% 0 0ms 0 0 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 6.3 5.8 66.7% 1 21.95s 7,152 187 5,882
Qwen3.6 35B A3B 10.0 10.0 100.0% 0 1.46s 7,788 248 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 19.00s 245 8 2,796
Qwen3.6 35B A3B 3.5 4.4 33.3% 2 7.45s 781 11,381 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.9 2.5 33.3% 1 38.70s 486 92 5,204
Qwen3.6 35B A3B 4.4 3.0 33.3% 1 3.51s 520 1,545 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.8 5.8 33.3% 1 12.80s 687 350 2,600
Qwen3.6 35B A3B 6.2 5.8 66.7% 1 1.86s 709 1,264 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 5.9 7.2 55.6% 1 24.87s 675 362 7,840
Qwen3.6 35B A3B 3.2 9.9 0.0% 0 1.07s 714 573 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 4.7 1.6 66.7% 1 12.05s 7,067 304 1,001
Qwen3.6 35B A3B 3.0 10.0 0.0% 0 0ms 0 0 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 3.0 10.0 0.0% 0 22.77s 211 3,068 3,452
Qwen3.6 35B A3B 3.0 10.0 0.0% 0 414ms 210 12 0

Quick Compare

Switch Comparison Pair