Navigate
AI BENCHY
Advertise here

Gemini 3.5 Flash (minimal) vs GPT-5.5 (medium)

GPT-5.5 (medium) leads on average score with 9.0 vs 6.8. Gemini 3.5 Flash (minimal) has the lower benchmark cost at $0.300 vs $4.137. Gemini 3.5 Flash (minimal) is faster at 2.65s vs 38.42s, with pass rates of 65.2% vs 87.9%.

Last updated at: 2026-07-25

Rank
#96
Total Output Tokens
16,454
Response Time (avg)
2.65s
Total Cost
$0.300
Rank
#15
Total Output Tokens
124,436
Response Time (avg)
38.42s
Total Cost
$4.137
Recommended model Gemini 3.5 Flash (minimal)

It offers the best overall trade-off: a competitive score (6.8), lower cost than GPT-5.5 (medium), and balanced response time.

Detailed comparison

Metric Gemini 3.5 Flash Gemini 3.5 Flash minimal Release: 2026-05-19 GPT-5.5 GPT-5.5 medium Release: 2026-04-24
Score 6.8 9.0
Rank #96 #15
Reliability 10.0 10.0
Consistency 9.6 8.9
Tests Correct
Attempt pass rate 65.2% 87.9%
Flaky tests 1 3
Total Runs 66 66
Cost per result 2.138 22.980
Total Cost $0.300 $4.137
Input Price $1.500 / 1M $5.000 / 1M
Output Price $9.000 / 1M $30.000 / 1M
Total Input Tokens 100,753 80,659
Output Tokens 16,454 5,617
Reasoning Tokens 0 118,819
Response Time (avg) 2.65s 38.42s
Response Time (max) 25.26s 332.10s
Response Time (total) 58.27s 845.35s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#96 Gemini 3.5 Flash

minimal
Cost
$0.041
Time
20.4s
Tokens
4,608 tok

#15 GPT-5.5

medium
Cost
$0.112
Time
71.9s
Tokens
3,807 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 5.6 9.9 33.3% 0 2.75s 8,122 3,456 0
GPT-5.5 8.8 7.8 88.9% 1 59.77s 7,305 362 24,959

Quick Compare

Switch Comparison Pair