Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Gemini 3.5 Flash (medium) vs Qwen3.8 Max (0902) (low)

Qwen3.8 Max (0902) (low) leads on average score with 9.1 vs 9.0. Qwen3.8 Max (0902) (low) has the lower benchmark cost at $0.664 vs $0.665. Gemini 3.5 Flash (medium) is faster at 8.48s vs 24.21s, with pass rates of 83.3% vs 86.4%.

Last updated at: 2026-09-07

Rank
#27
Total Output Tokens
62,173
Response Time (avg)
8.48s
Total Cost
$0.665
Rank
#23
Total Output Tokens
75,865
Response Time (avg)
24.21s
Total Cost
$0.664
Recommended model Gemini 3.5 Flash (medium)

Its score stays close to the best score here (9.0 vs 9.1), while responding about 2.9x faster than Qwen3.8 Max (0902) (low).

Detailed comparison

Metric Gemini 3.5 Flash Gemini 3.5 Flash medium Release: 2026-05-19 Qwen3.8 Max (0902) Qwen3.8 Max (0902) low Release: 2026-09-07
Score 9.0 9.1
Rank #27 #23
Reliability 10.0 10.0
Consistency 9.7 9.3
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 83.3% 86.4%
Flaky tests 1 2
Total Runs 66 66
Cost per result 3.690 3.684
Total Cost $0.665 $0.664
Input Price $1.500 / 1M $2.000 / 1M
Output Price $9.000 / 1M $6.000 / 1M
Total Input Tokens 69,756 103,954
Output Tokens 2,166 6,322
Reasoning Tokens 60,007 69,543
Response Time (avg) 8.48s 24.21s
Response Time (max) 76.68s 207.79s
Response Time (total) 186.57s 532.62s
Parameters ~500B total (~20B active) 2.4T total (~100B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#27 Gemini 3.5 Flash

medium
Cost
$0.201
Time
112.9s
Tokens
22,371 tok

#23 Qwen3.8 Max (0902)

low
Cost
$0.014
Time
41.2s
Tokens
2,421 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 7.9 7.5 77.8% 1 12.63s 8,118 461 24,939
Qwen3.8 Max (0902) 10.0 10.0 100.0% 0 37.42s 8,127 485 17,404

Quick Compare

Switch Comparison Pair