Navigate
AI BENCHY
Advertise here

Seed-2.0-Code — high vs medium

The average score is effectively tied at 6.9 vs 6.9. medium has the lower benchmark cost at $1.049 vs $1.312. medium is faster at 93.60s vs 127.79s, with pass rates of 71.2% vs 71.2%.

Last updated at: 2026-08-12

Rank
#123
Total Output Tokens
423,424
Response Time (avg)
127.79s
Total Cost
$1.312
Rank
#124
Total Output Tokens
336,696
Response Time (avg)
93.60s
Total Cost
$1.049
Recommended model medium

It has the strongest score in this comparison (6.9) and the best overall balance of cost and response time across all 2 models.

Detailed comparison

Metric Seed-2.0-Code Seed-2.0-Code high Release: 2026-08-12 Seed-2.0-Code Seed-2.0-Code medium Release: 2026-08-12
Score 6.9 6.9
Rank #123 #124
Reliability 6.0 4.7
Consistency 7.0 6.7
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 71.2% 71.2%
Flaky tests 8 9
Total Runs 66 66
Cost per result 11.921 10.482
Total Cost $1.312 $1.049
Input Price $0.500 / 1M $0.500 / 1M
Output Price $3.000 / 1M $3.000 / 1M
Total Input Tokens 81,961 76,027
Output Tokens 7,187 7,268
Reasoning Tokens 416,237 329,428
Response Time (avg) 127.79s 93.60s
Response Time (max) 675.30s 626.58s
Response Time (total) 2811.27s 2059.27s
Parameters ~200B total (~20B active) ~200B total (~20B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#123 Seed-2.0-Code

high
Cost
$0.037
Time
183.6s
Tokens
12,388 tok

#124 Seed-2.0-Code

medium
Cost
$0.026
Time
122.8s
Tokens
8,630 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Code 5.9 4.4 66.7% 2 266.74s 6,750 432 124,956
Seed-2.0-Code 8.9 7.8 88.9% 1 89.94s 8,247 496 44,062

Quick Compare

Switch Comparison Pair