Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

DeepSeek: DeepSeek V4 Flash vs Google: Gemini 3 Flash Preview

Last updated at: 2026-05-26

Metric DeepSeek V4 Flash DeepSeek V4 Flash high Release: 2026-04-24 Free Available Gemini 3 Flash Preview Gemini 3 Flash Preview none Release: 2025-12-17
Score 7.6 7.7
Rank #44 #41
Reliability 10.0 10.0
Consistency 8.4 9.2
Tests Correct
Attempt pass rate 73.3% 70.0%
Flaky tests 4 2
Total Runs 98 98
Cost per result 0.329 0.196
Total Cost $0.040 $0.026
Input Price $0.112 / 1M $0.500 / 1M
Output Price $0.224 / 1M $3.000 / 1M
Output Tokens 11,480 2,449
Reasoning Tokens 122,086 0
Response Time (avg) 46.36s 1.70s
Response Time (max) 218.13s 3.56s
Response Time (total) 927.27s 22.05s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 8.3 10.0 75.0% 0 28.51s 140 7,770
Gemini 3 Flash Preview 8.3 10.0 75.0% 0 1.25s 214 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 6.8 10.0 50.0% 0 58.13s 387 27,101
Gemini 3 Flash Preview 6.8 10.0 50.0% 0 2.19s 447 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 76.57s 465 7,347
Gemini 3 Flash Preview 4.7 1.6 66.7% 1 3.56s 350 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 28.03s 201 1,179
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 1.41s 279 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.1 4.4 44.5% 2 100.31s 27 59,249
Gemini 3 Flash Preview 7.7 10.0 66.7% 0 963ms 18 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 6.1 3.1 66.7% 1 25.15s 79 632
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 1.13s 104 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 15.36s 63 1,622
Gemini 3 Flash Preview 6.4 5.8 66.7% 1 1.58s 74 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 8.2 7.2 88.9% 1 26.11s 1,374 8,113
Gemini 3 Flash Preview 7.7 10.0 66.7% 0 1.05s 714 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 74.73s 228 542
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 3.35s 234 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.0 10.0 0.0% 0 54.46s 8,516 8,531
Gemini 3 Flash Preview 3.0 10.0 0.0% 0 1.07s 15 0

Quick Compare

Switch Comparison Pair