Navigate
AI BENCHY
Advertise here

Gemini 3.5 Flash vs Qwen3.7 Plus (medium)

Qwen3.7 Plus (medium) leads on average score with 7.9 vs 7.0. Qwen3.7 Plus (medium) has the lower benchmark cost at $0.267 vs $1.079. Gemini 3.5 Flash is faster at 9.93s vs 51.51s, with pass rates of 74.2% vs 75.8%.

Last updated at: 2026-07-25

Rank
#87
Total Output Tokens
117,518
Response Time (avg)
9.93s
Total Cost
$1.079
Rank
#43
Total Output Tokens
179,429
Response Time (avg)
51.51s
Total Cost
$0.267
Recommended model Qwen3.7 Plus (medium)

It has the best score here (7.9), while costing about 4.0x less than Gemini 3.5 Flash.

Detailed comparison

Metric Gemini 3.5 Flash Gemini 3.5 Flash none Release: 2026-05-19 Qwen3.7 Plus Qwen3.7 Plus medium Release: 2026-06-03
Score 7.0 7.9
Rank #87 #43
Reliability 10.0 10.0
Consistency 8.9 8.9
Tests Correct
Attempt pass rate 74.2% 75.8%
Flaky tests 3 3
Total Runs 66 66
Cost per result 7.190 2.072
Total Cost $1.079 $0.267
Input Price $1.500 / 1M $0.320 / 1M
Output Price $9.000 / 1M $1.280 / 1M
Total Input Tokens 13,843 115,233
Output Tokens 117,518 6,162
Reasoning Tokens 0 173,267
Response Time (avg) 9.93s 51.51s
Response Time (max) 64.36s 315.30s
Response Time (total) 178.68s 1133.15s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#87 Gemini 3.5 Flash

none
Cost
$0.225
Time
125.5s
Tokens
25,004 tok

#43 Qwen3.7 Plus

medium
Cost
$0.018
Time
193.2s
Tokens
10,821 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 8.8 7.8 88.9% 1 34.69s 8,122 75,927 0
Qwen3.7 Plus 6.1 6.6 55.6% 1 108.60s 6,472 414 43,576

Quick Compare

Switch Comparison Pair