Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

ByteDance Seed: Seed-2.0-Mini vs Qwen: Qwen3.6 27B

Summary

Seed-2.0-Mini vs Qwen3.6 27B benchmark comparison: Seed-2.0-Mini leads on average score with 6.9 vs 6.8. Seed-2.0-Mini has the lower benchmark cost at $0.044 vs $0.336. Qwen3.6 27B is faster at 59.71s vs 80.22s, with pass rates of 57.1% vs 60.3%.

Recommended model: Seed-2.0-Mini - It has the best score here (6.9), while costing about 7.7x less than Qwen3.6 27B.

Last updated at: 2026-06-10

Metric Seed-2.0-Mini Seed-2.0-Mini medium Release: 2026-02-14 Qwen3.6 27B Qwen3.6 27B medium Release: 2026-04-20
Score 6.9 6.8
Rank #74 #79
Reliability 6.7 10.0
Consistency 9.3 8.2
Tests Correct
Attempt pass rate 57.1% 60.3%
Flaky tests 2 5
Total Runs 63 63
Cost per result 0.397 3.361
Total Cost $0.044 $0.336
Input Price $0.100 / 1M $0.290 / 1M
Output Price $0.400 / 1M $2.400 / 1M
Total Input Tokens 41,904 39,376
Output Tokens 2,555 16,189
Reasoning Tokens 95,974 122,521
Response Time (avg) 80.22s 59.71s
Response Time (max) 262.83s 168.22s
Response Time (total) 1363.72s 1254.01s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#74 Seed-2.0-Mini

medium
Cost
$0.002
Time
161.7s
Tokens
4,379 tok

#79 Qwen3.6 27B

medium
Cost
$0.009
Time
39.6s
Tokens
3,090 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 6.6 10.0 50.0% 0 74.75s 791 360 9,520
Qwen3.6 27B 8.3 10.0 75.0% 0 12.62s 453 582 4,311
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 5.5 9.8 33.3% 0 220.48s 3,823 464 34,964
Qwen3.6 27B 7.7 10.0 66.7% 0 142.99s 5,051 7,968 43,367
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 10.0 10.0 100.0% 0 262.83s 16,533 404 29,806
Qwen3.6 27B 7.0 3.7 66.7% 1 83.07s 15,104 2,088 14,689
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 10.0 10.0 100.0% 0 24.27s 8,568 246 2,743
Qwen3.6 27B 3.5 1.4 50.0% 2 37.30s 7,778 568 9,404
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 3.0 10.0 0.0% 0 0ms 0 0 0
Qwen3.6 27B 2.9 7.2 11.1% 1 73.38s 662 3,510 20,352
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 5.1 3.4 33.3% 1 36.65s 585 213 4,210
Qwen3.6 27B 6.5 3.4 66.7% 1 39.53s 516 81 3,045
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 10.0 10.0 100.0% 0 17.47s 840 69 2,050
Qwen3.6 27B 10.0 10.0 100.0% 0 37.96s 699 346 6,548
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 8.2 7.2 88.9% 1 31.79s 903 527 5,667
Qwen3.6 27B 7.7 10.0 66.7% 0 61.14s 696 255 12,044
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 10.0 10.0 100.0% 0 88.68s 9,585 222 5,235
Qwen3.6 27B 10.0 10.0 100.0% 0 16.88s 8,213 390 2,954
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Mini 3.0 10.0 0.0% 0 56.76s 276 50 1,779
Qwen3.6 27B 3.0 10.0 0.0% 0 80.99s 204 401 5,807

Quick Compare

Switch Comparison Pair