Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

ByteDance Seed: Seed-2.0-Lite vs Google: Gemini 3 Flash Preview

Last updated at: 2026-03-12

Metric Seed-2.0-Lite Seed-2.0-Lite medium Release: 2026-02-14 Gemini 3 Flash Preview Gemini 3 Flash Preview none Release: 2025-12-17
Rank #3 #21
Avg Score 8.5 7.2
Consistency 8.7 9.0
Cost per result 0.870 0.169
Total Cost $0.105 $0.019
Tests Correct
Attempt pass rate 87.5% 75.0%
Flaky tests 3 2
Total Runs 48 48
Output Tokens 2,815 1,411
Reasoning Tokens 44,618 0
Response Time (avg) 29.39s 1.75s
Response Time (max) 168.71s 3.56s
Response Time (total) 470.29s 15.71s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Avg Score vs Response Time (avg)

Total Output Tokens

Avg Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 23.34s 990 7,037
Gemini 3 Flash Preview 7.0 10.0 66.7% 0 1.59s 208 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 37.67s 506 4,299
Gemini 3 Flash Preview 10.0 1.6 66.7% 1 3.56s 350 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 9.9 10.0 100.0% 0 9.07s 246 1,742
Gemini 3 Flash Preview 9.9 10.0 100.0% 0 1.41s 279 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 4.0 7.2 55.6% 1 88.74s 15 23,897
Gemini 3 Flash Preview 7.0 10.0 66.7% 0 963ms 18 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 7.0 3.6 66.7% 1 18.25s 304 1,620
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 1.13s 104 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 7.26s 71 1,480
Gemini 3 Flash Preview 5.5 5.8 66.7% 1 1.58s 74 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 9.3 7.9 88.9% 1 11.03s 461 3,532
Gemini 3 Flash Preview 7.0 10.0 66.7% 0 1.06s 144 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Seed-2.0-Lite 10.0 10.0 100.0% 0 12.38s 222 1,011
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 3.35s 234 0

Quick Compare

Switch Comparison Pair