Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

ByteDance Seed: Seed-2.0-Lite vs DeepSeek: DeepSeek V3.2

Last updated at: 2026-06-01

Metric Seed-2.0-Lite Seed-2.0-Lite none Release: 2026-02-14 DeepSeek V3.2 DeepSeek V3.2 none Release: 2025-12-01
Score 5.9 5.6
Rank #106 #120
Reliability 10.0 10.0
Consistency 8.3 8.3
Tests Correct
Attempt pass rate 48.3% 41.7%
Flaky tests 4 6
Total Runs 60 60
Cost per result 0.218 0.222
Total Cost $0.018 $0.018
Input Price $0.250 / 1M $0.252 / 1M
Output Price $2.000 / 1M $0.378 / 1M
Output Tokens 3,253 11,159
Reasoning Tokens 0 0
Response Time (avg) 2.48s 14.43s
Response Time (max) 6.70s 115.89s
Response Time (total) 49.67s 288.55s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 3.0 5.9 16.7% 2 2.43s 709 0
DeepSeek V3.2 3.2 8.2 8.3% 1 9.35s 1,073 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 6.8 9.9 50.0% 0 2.95s 404 0
DeepSeek V3.2 3.1 5.4 16.7% 1 20.87s 4,522 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 3.0 10.0 0.0% 0 6.59s 498 0
DeepSeek V3.2 6.5 10.0 0.0% 0 115.89s 2,887 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 1.82s 246 0
DeepSeek V3.2 6.3 5.8 66.7% 1 9.42s 1,710 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 3.6 7.2 22.2% 1 1.33s 17 0
DeepSeek V3.2 2.9 6.9 11.1% 1 4.17s 21 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 3.45s 294 0
DeepSeek V3.2 6.8 10.0 66.7% 1 9.32s 43 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 1.06s 73 0
DeepSeek V3.2 10.0 10.0 100.0% 0 1.52s 66 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 5.3 7.2 44.4% 1 2.78s 709 0
DeepSeek V3.2 8.3 10.0 77.8% 1 6.91s 298 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 3.94s 292 0
DeepSeek V3.2 10.0 10.0 100.0% 0 11.85s 522 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 3.0 10.0 0.0% 0 1.96s 11 0
DeepSeek V3.2 3.0 10.0 0.0% 0 17.23s 17 0

Quick Compare

Switch Comparison Pair