Navigate
AI BENCHY
Advertise here

Seed-2.0-Code vs Laguna XS 2.1

Laguna XS 2.1 leads on average score with 5.3 vs 5.3. Laguna XS 2.1 has the lower benchmark cost at $0.008 vs $0.151. Laguna XS 2.1 is faster at 1.55s vs 14.67s, with pass rates of 34.9% vs 30.3%.

Last updated at: 2026-08-12

Rank
#219
Total Output Tokens
23,432
Response Time (avg)
14.67s
Total Cost
$0.151
Rank
#215
Total Output Tokens
13,377
Response Time (avg)
1.55s
Total Cost
$0.008
Recommended model Laguna XS 2.1

It has the best score here (5.3), while costing about 21.1x less than Seed-2.0-Code.

Detailed comparison

Metric Seed-2.0-Code Seed-2.0-Code none Release: 2026-08-12 Laguna XS 2.1 Laguna XS 2.1 none Release: 2026-07-02 Free Available
Score 5.3 5.3
Rank #219 #215
Reliability 6.7 10.0
Consistency 8.1 9.0
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 34.9% 30.3%
Flaky tests 5 3
Total Runs 66 66
Cost per result 2.503 0.143
Total Cost $0.151 $0.008
Input Price $0.500 / 1M $0.060 / 1M
Output Price $3.000 / 1M $0.120 / 1M
Total Input Tokens 159,740 91,598
Output Tokens 23,432 13,377
Reasoning Tokens 0 0
Response Time (avg) 14.67s 1.55s
Response Time (max) 132.84s 19.02s
Response Time (total) 322.75s 34.19s
Parameters ~200B total (~20B active) 33B total (3B active)
Availability Closed Weights available

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#219 Seed-2.0-Code

none
Cost
$0.030
Time
140.6s
Tokens
9,924 tok

#215 Laguna XS 2.1

none
Cost
$0.001
Time
27.6s
Tokens
4,344 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Code 3.6 9.8 0.0% 0 13.42s 7,230 505 0
Laguna XS 2.1 4.3 7.8 22.2% 1 623ms 7,995 562 0

Quick Compare

Switch Comparison Pair