Navigate
AI BENCHY
Advertise here

Gemini 3.1 Flash Lite vs Qwen3.5-35B-A3B

The average score is effectively tied at 6.1 vs 6.1. Gemini 3.1 Flash Lite has the lower benchmark cost at $0.046 vs $0.106. Gemini 3.1 Flash Lite is faster at 1.75s vs 12.72s, with pass rates of 50.0% vs 43.9%.

Last updated at: 2026-08-05

Rank
#147
Total Output Tokens
10,723
Response Time (avg)
1.75s
Total Cost
$0.046
Rank
#152
Total Output Tokens
86,614
Response Time (avg)
12.72s
Total Cost
$0.106
Recommended model Gemini 3.1 Flash Lite

It has the best score here (6.1), while costing about 2.3x less than Qwen3.5-35B-A3B.

Detailed comparison

Metric Gemini 3.1 Flash Lite Gemini 3.1 Flash Lite none Release: 2026-05-08 Qwen3.5-35B-A3B Qwen3.5-35B-A3B none Release: 2026-02-24
Score 6.1 6.1
Rank #147 #152
Reliability 10.0 10.0
Consistency 8.6 8.6
Benchmark coverage 22/22 tests · 66/66 attempts 22/22 tests · 66/66 attempts
Tests Correct
Attempt pass rate 50.0% 43.9%
Flaky tests 4 4
Total Runs 66 66
Cost per result 0.507 1.578
Total Cost $0.046 $0.106
Input Price $0.250 / 1M $0.140 / 1M
Output Price $1.500 / 1M $1.000 / 1M
Total Input Tokens 118,050 134,521
Output Tokens 10,723 86,614
Reasoning Tokens 0 0
Response Time (avg) 1.75s 12.72s
Response Time (max) 16.25s 209.15s
Response Time (total) 38.60s 279.90s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#147 Gemini 3.1 Flash Lite

none
Cost
$0.001
Time
4.5s
Tokens
727 tok

#152 Qwen3.5-35B-A3B

none
Cost
$0.005
Time
28.4s
Tokens
4,518 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 5.5 10.0 33.3% 0 938ms 8,128 666 0
Qwen3.5-35B-A3B 5.5 10.0 33.3% 0 1.39s 7,808 571 0

Quick Compare

Switch Comparison Pair