Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Google: Gemini 3.1 Flash Lite vs Qwen: Qwen3.5-122B-A10B

Summary

Gemini 3.1 Flash Lite vs Qwen3.5-122B-A10B benchmark comparison: The average score is effectively tied at 7.8 vs 7.7. Gemini 3.1 Flash Lite has the lower benchmark cost at $0.071 vs $0.588. Gemini 3.1 Flash Lite is faster at 3.23s vs 42.49s, with pass rates of 65.1% vs 73.0%.

Recommended model: Gemini 3.1 Flash Lite - It has the best score here (7.8), while costing about 8.4x less than Qwen3.5-122B-A10B.

Last updated at: 2026-06-18

Metric Gemini 3.1 Flash Lite Gemini 3.1 Flash Lite medium Release: 2026-05-08 Qwen3.5-122B-A10B Qwen3.5-122B-A10B medium Release: 2026-02-24
Score 7.8 7.7
Rank #34 #36
Reliability 10.0 10.0
Consistency 9.2 8.8
Tests Correct
Attempt pass rate 65.1% 73.0%
Flaky tests 2 3
Total Runs 63 63
Cost per result 0.539 5.235
Total Cost $0.071 $0.588
Input Price $0.250 / 1M $0.260 / 1M
Output Price $1.500 / 1M $2.080 / 1M
Total Input Tokens 36,808 41,832
Output Tokens 2,254 26,187
Reasoning Tokens 38,300 251,028
Response Time (avg) 3.23s 42.49s
Response Time (max) 10.87s 168.16s
Response Time (total) 67.80s 892.30s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#34 Gemini 3.1 Flash Lite

medium
Cost
$0.003
Time
5.3s
Tokens
1,754 tok

#36 Qwen3.5-122B-A10B

medium
Cost
$0.019
Time
48.7s
Tokens
6,034 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 9.1 10.0 75.0% 0 2.39s 502 604 4,201
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 9.75s 672 269 16,835
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 5.5 10.0 33.3% 0 3.81s 8,134 459 8,978
Qwen3.5-122B-A10B 6.0 7.2 55.6% 1 114.48s 7,630 8,057 82,578
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 10.0 10.0 100.0% 0 10.87s 12,873 327 7,401
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 107.79s 14,947 483 11,337
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 10.0 10.0 100.0% 0 2.60s 7,362 279 2,845
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 23.41s 7,782 270 16,558
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 2.9 7.2 11.1% 1 3.16s 643 15 5,165
Qwen3.5-122B-A10B 2.9 7.2 11.1% 1 63.40s 771 15,537 64,889
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 10.0 10.0 100.0% 0 2.60s 488 84 1,142
Qwen3.5-122B-A10B 3.4 2.2 33.3% 1 34.11s 344 66 7,592
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 9.9 10.0 100.0% 0 2.59s 623 75 3,320
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 9.88s 593 77 7,372
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 7.6 7.2 77.8% 1 1.95s 568 165 2,450
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 17.89s 696 284 27,575
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 10.0 10.0 100.0% 0 4.55s 5,457 234 921
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 4.60s 8,193 322 1,226
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 3.0 10.0 0.0% 0 3.08s 158 12 1,877
Qwen3.5-122B-A10B 3.0 10.0 0.0% 0 52.87s 204 822 15,066

Quick Compare

Switch Comparison Pair