Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Google: Gemini 3.5 Flash vs Qwen: Qwen3.6 35B A3B

Summary

Gemini 3.5 Flash vs Qwen3.6 35B A3B benchmark comparison: Gemini 3.5 Flash leads on average score with 6.8 vs 6.7. Gemini 3.5 Flash has the lower benchmark cost at $0.108 vs $0.146. Gemini 3.5 Flash is faster at 1.57s vs 18.08s, with pass rates of 68.3% vs 63.5%.

Recommended model: Gemini 3.5 Flash - It has the best score here (6.8), while responding about 11.5x faster than Qwen3.6 35B A3B.

Last updated at: 2026-07-02

Metric Gemini 3.5 Flash Gemini 3.5 Flash minimal Release: 2026-05-19 Qwen3.6 35B A3B Qwen3.6 35B A3B medium Release: 2026-04-20
Score 6.8 6.7
Rank #74 #78
Reliability 10.0 10.0
Consistency 9.6 9.6
Tests Correct
Attempt pass rate 68.3% 63.5%
Flaky tests 1 1
Total Runs 63 63
Cost per result 0.767 1.094
Total Cost $0.108 $0.146
Input Price $1.500 / 1M $0.140 / 1M
Output Price $9.000 / 1M $1.000 / 1M
Total Input Tokens 39,847 16,385
Output Tokens 5,277 19,632
Reasoning Tokens 0 130,219
Response Time (avg) 1.57s 18.08s
Response Time (max) 5.51s 86.11s
Response Time (total) 33.02s 343.61s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#74 Gemini 3.5 Flash

minimal
Cost
$0.041
Time
20.4s
Tokens
4,608 tok

#78 Qwen3.6 35B A3B

medium
Invalid SVG
Cost
$0.000
Time
300.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 6.5 10.0 50.0% 0 892ms 492 405 0
Qwen3.6 35B A3B 10.0 10.0 100.0% 0 6.02s 672 1,154 12,385
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 5.6 9.9 33.3% 0 2.75s 8,122 3,456 0
Qwen3.6 35B A3B 7.7 10.0 66.7% 0 50.55s 5,051 7,929 37,223
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 3.0 10.0 0.0% 0 3.56s 15,780 404 0
Qwen3.6 35B A3B 3.0 10.0 0.0% 0 0ms 0 0 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 1.66s 7,548 279 0
Qwen3.6 35B A3B 10.0 10.0 100.0% 0 12.99s 7,776 2,591 9,968
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 899ms 633 12 0
Qwen3.6 35B A3B 5.3 7.2 44.4% 1 22.50s 771 6,193 39,116
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 922ms 486 117 0
Qwen3.6 35B A3B 4.4 9.9 0.0% 0 8.66s 516 129 4,569
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 6.4 5.8 66.7% 1 893ms 615 76 0
Qwen3.6 35B A3B 10.0 10.0 100.0% 0 7.50s 699 219 7,404
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 1.45s 558 282 0
Qwen3.6 35B A3B 8.0 10.0 66.7% 0 5.95s 696 655 9,228
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 2.79s 5,457 234 0
Qwen3.6 35B A3B 3.0 10.0 0.0% 0 0ms 0 0 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 3.0 10.0 0.0% 0 1.76s 156 12 0
Qwen3.6 35B A3B 3.0 10.0 0.0% 0 32.90s 204 762 10,326

Quick Compare

Switch Comparison Pair