Navigate
AI BENCHY
Advertise here

Laguna S 2.1 (low) vs MiMo-V2.5

MiMo-V2.5 leads on average score with 5.1 vs 5.0. MiMo-V2.5 has the lower benchmark cost at $0.025 vs $0.082. MiMo-V2.5 is faster at 4.68s vs 85.34s, with pass rates of 22.7% vs 27.3%.

Last updated at: 2026-09-04

Rank
#266
Total Output Tokens
451,708
Response Time (avg)
85.34s
Total Cost
$0.082
Rank
#260
Total Output Tokens
16,464
Response Time (avg)
4.68s
Total Cost
$0.025
Recommended model MiMo-V2.5

It has the best score here (5.1), while costing about 3.3x less than Laguna S 2.1 (low).

Detailed comparison

Metric Laguna S 2.1 Laguna S 2.1 low Release: 2026-07-21 Free Available MiMo-V2.5 MiMo-V2.5 none Release: 2026-04-22
Score 5.0 5.1
Rank #266 #260
Reliability 10.0 10.0
Consistency 8.0 9.2
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 22.7% 27.3%
Flaky tests 5 2
Total Runs 66 66
Cost per result 3.022 0.769
Total Cost $0.082 $0.025
Input Price $0.090 / 1M $0.140 / 1M
Output Price $0.180 / 1M $0.280 / 1M
Total Input Tokens 118,752 141,052
Output Tokens 67,823 16,464
Reasoning Tokens 383,885 0
Response Time (avg) 85.34s 4.68s
Response Time (max) 814.73s 55.36s
Response Time (total) 1877.49s 103.02s
Parameters 118B total (8B active) 310B total (15B active)
Availability Weights available Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#266 Laguna S 2.1

low
Cost
$0.001
Time
6.6s
Tokens
1,314 tok

#260 MiMo-V2.5

none
Cost
$0.007
Time
267.4s
Tokens
25,283 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Laguna S 2.1 3.8 7.1 22.2% 1 167.51s 7,917 58,111 119,138
MiMo-V2.5 5.5 10.0 33.3% 0 3.24s 7,440 696 0

Quick Compare

Switch Comparison Pair