Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Google: Gemini 3.5 Flash vs OpenAI: GPT-5.5

Last updated at: 2026-05-19

Metric Gemini 3.5 Flash Gemini 3.5 Flash high Release: 2026-05-19 GPT-5.5 GPT-5.5 medium Release: 2026-04-24
Score 9.6 8.9
Rank #4 #8
Reliability 10.0 10.0
Consistency 9.6 9.1
Tests Correct
Attempt pass rate 96.5% 87.7%
Flaky tests 1 2
Total Runs 57 57
Cost per result 4.294 18.365
Total Cost $0.773 $2.939
Input Price $1.500 / 1M $5.000 / 1M
Output Price $9.000 / 1M $30.000 / 1M
Output Tokens 1,945 1,950
Reasoning Tokens 78,877 91,386
Response Time (avg) 6.90s 33.02s
Response Time (max) 22.37s 332.10s
Response Time (total) 131.10s 627.45s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 2.57s 174 4,997
GPT-5.5 10.0 10.0 100.0% 0 4.66s 250 1,335
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 14.42s 426 10,368
GPT-5.5 10.0 10.0 100.0% 0 9.09s 318 1,391
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 22.37s 351 16,323
GPT-5.5 10.0 10.0 100.0% 0 19.29s 312 2,841
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 6.43s 279 8,466
GPT-5.5 10.0 10.0 100.0% 0 4.18s 234 593
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 7.6 7.2 77.8% 1 14.09s 12 24,721
GPT-5.5 5.3 7.2 44.4% 1 164.14s 67 79,625
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 3.63s 115 1,650
GPT-5.5 10.0 10.0 100.0% 0 4.16s 138 223
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 3.35s 70 3,799
GPT-5.5 10.0 10.0 100.0% 0 3.36s 93 538
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 3.23s 241 4,940
GPT-5.5 10.0 10.0 100.0% 0 6.78s 250 2,254
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 9.8 10.0 100.0% 0 4.96s 265 1,608
GPT-5.5 10.0 10.0 100.0% 0 10.57s 258 832
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3.5 Flash 10.0 10.0 100.0% 0 3.94s 12 2,005
GPT-5.5 2.8 1.6 33.3% 1 37.86s 30 1,754

Quick Compare

Switch Comparison Pair