Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

MiniMax: MiniMax M2.7 vs Poolside: Laguna S 2.1

MiniMax M2.7 (medium) leads on average score with 5.0 vs 4.5. Laguna S 2.1 has the lower benchmark cost at $0.025 vs $0.163. Laguna S 2.1 is faster at 11.67s vs 41.28s, with pass rates of 45.5% vs 15.2%.

Recommended modelLaguna S 2.1Its score stays close to the best score here (4.5 vs 5.0), while costing about 6.8x less than MiniMax M2.7 (medium).

Last updated at: 2026-07-22

Metric MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18 Laguna S 2.1 Laguna S 2.1 none Release: 2026-07-21 Free Available
Score 5.0 4.5
Rank #180 #200
Reliability 10.0 10.0
Consistency 6.6 8.8
Tests Correct
Attempt pass rate 45.5% 15.2%
Flaky tests 9 3
Total Runs 66 66
Cost per result 3.906 1.205
Total Cost $0.163 $0.025
Input Price $0.250 / 1M $0.100 / 1M
Output Price $1.000 / 1M $0.200 / 1M
Total Input Tokens 114,518 125,598
Output Tokens 18,558 57,634
Reasoning Tokens 119,036 0
Response Time (avg) 41.28s 11.67s
Response Time (max) 196.21s 200.52s
Response Time (total) 866.81s 256.80s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#180 MiniMax M2.7

medium
Cost
$0.022
Time
22.8s
Tokens
9,250 tok

#200 Laguna S 2.1

none
Cost
$0.001
Time
7.6s
Tokens
1,345 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
MiniMax M2.7 5.7 9.1 33.3% 0 101.89s 2,961 1,231 38,841
Laguna S 2.1 3.4 9.7 0.0% 0 550ms 7,917 424 0

Quick Compare

Switch Comparison Pair