Navigate
AI BENCHY
Advertise here

Seed-2.0-Code vs MiniMax M2.7 (medium)

Seed-2.0-Code leads on average score with 5.3 vs 5.0. Seed-2.0-Code has the lower benchmark cost at $0.151 vs $0.196. Seed-2.0-Code is faster at 14.67s vs 41.28s, with pass rates of 34.9% vs 45.5%.

Last updated at: 2026-08-12

Rank
#219
Total Output Tokens
23,432
Response Time (avg)
14.67s
Total Cost
$0.151
Rank
#229
Total Output Tokens
137,594
Response Time (avg)
41.28s
Total Cost
$0.196
Recommended model Seed-2.0-Code

It has the best score here (5.3), while responding about 2.8x faster than MiniMax M2.7 (medium).

Detailed comparison

Metric Seed-2.0-Code Seed-2.0-Code none Release: 2026-08-12 MiniMax M2.7 MiniMax M2.7 medium Release: 2026-03-18
Score 5.3 5.0
Rank #219 #229
Reliability 6.7 10.0
Consistency 8.1 6.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 34.9% 45.5%
Flaky tests 5 9
Total Runs 66 66
Cost per result 2.503 3.906
Total Cost $0.151 $0.196
Input Price $0.500 / 1M $0.300 / 1M
Output Price $3.000 / 1M $1.200 / 1M
Total Input Tokens 159,740 114,518
Output Tokens 23,432 18,558
Reasoning Tokens 0 119,036
Response Time (avg) 14.67s 41.28s
Response Time (max) 132.84s 196.21s
Response Time (total) 322.75s 866.81s
Parameters ~200B total (~20B active) 230B total (10B active)
Availability Closed Weights available

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#219 Seed-2.0-Code

none
Cost
$0.030
Time
140.6s
Tokens
9,924 tok

#229 MiniMax M2.7

medium
Cost
$0.022
Time
22.8s
Tokens
9,250 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Code 3.6 9.8 0.0% 0 13.42s 7,230 505 0
MiniMax M2.7 5.7 9.1 33.3% 0 101.89s 2,961 1,231 38,841

Quick Compare

Switch Comparison Pair