Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

ByteDance Seed: Seed-2.0-Lite vs Google: Gemini 3.1 Flash Lite Preview

Summary

Seed-2.0-Lite vs Gemini 3.1 Flash Lite Preview benchmark comparison: Seed-2.0-Lite leads on average score with 8.2 vs 7.2. Gemini 3.1 Flash Lite Preview has the lower benchmark cost at $0.018 vs $0.175. Gemini 3.1 Flash Lite Preview is faster at 1.21s vs 47.07s, with pass rates of 76.2% vs 60.3%.

Recommended model: Gemini 3.1 Flash Lite Preview - It offers the best overall trade-off: a competitive score (7.2), lower cost than Seed-2.0-Lite, and balanced response time.

Last updated at: 2026-06-10

Metric Seed-2.0-Lite Seed-2.0-Lite medium Release: 2026-02-14 Gemini 3.1 Flash Lite Preview Gemini 3.1 Flash Lite Preview none Release: 2026-03-03
Score 8.2 7.2
Rank #20 #59
Reliability 10.0 10.0
Consistency 9.0 9.7
Tests Correct
Attempt pass rate 76.2% 60.3%
Flaky tests 3 1
Total Runs 63 63
Cost per result 1.250 0.148
Total Cost $0.175 $0.018
Input Price $0.250 / 1M $0.250 / 1M
Output Price $2.000 / 1M $1.500 / 1M
Total Input Tokens 46,740 37,582
Output Tokens 3,230 5,547
Reasoning Tokens 78,406 0
Response Time (avg) 47.07s 1.21s
Response Time (max) 254.92s 3.39s
Response Time (total) 988.37s 25.45s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#20 Seed-2.0-Lite

medium
Cost
$0.005
Time
86.7s
Tokens
2,354 tok

#59 Gemini 3.1 Flash Lite Preview

none
Cost
$0.003
Time
4.7s
Tokens
1,827 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 8.3 10.0 75.0% 0 17.99s 942 996 7,142
Gemini 3.1 Flash Lite Preview 7.5 8.4 66.7% 1 1.04s 504 1,092 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 8.0 9.8 66.7% 0 156.74s 8,247 458 31,890
Gemini 3.1 Flash Lite Preview 5.5 10.0 33.3% 0 967ms 8,128 670 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 37.67s 16,254 506 4,299
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 3.20s 13,026 339 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 9.07s 8,562 246 1,742
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.22s 7,550 399 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 5.9 7.2 55.6% 1 88.74s 843 15 23,897
Gemini 3.1 Flash Lite Preview 5.3 10.0 33.3% 0 942ms 641 568 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 6.7 3.6 66.7% 1 18.25s 582 304 1,620
Gemini 3.1 Flash Lite Preview 4.0 10.0 0.0% 0 741ms 488 69 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 7.26s 834 71 1,480
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.13s 623 574 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 9.0 7.9 88.9% 1 10.23s 894 403 3,285
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 900ms 570 1,045 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 12.38s 9,306 222 1,011
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 3.39s 5,894 782 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Seed-2.0-Lite 3.0 10.0 0.0% 0 48.32s 276 9 2,040
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 814ms 158 9 0

Quick Compare

Switch Comparison Pair