Navigate
AI BENCHY
Advertise here

Google: Gemini 3.5 Flash vs Thinking Machines: Inkling

Gemini 3.5 Flash (medium) leads on average score with 9.1 vs 8.0. Gemini 3.5 Flash (medium) has the lower benchmark cost at $0.642 vs $1.006. Gemini 3.5 Flash (medium) is faster at 8.20s vs 64.16s, with pass rates of 87.9% vs 77.3%.

Recommended modelGemini 3.5 Flash (medium)It has the best score here (9.1), while costing about 1.6x less than Inkling (high).

Last updated at: 2026-07-21

Metric Gemini 3.5 Flash Gemini 3.5 Flash medium Release: 2026-05-19 Inkling Inkling high Release: 2026-07-18
Score 9.1 8.0
Rank #12 #32
Reliability 10.0 9.8
Consistency 9.7 8.9
Tests Correct
Attempt pass rate 87.9% 77.3%
Flaky tests 1 3
Total Runs 66 66
Cost per result 3.374 6.704
Total Cost $0.642 $1.006
Input Price $1.500 / 1M $1.000 / 1M
Output Price $9.000 / 1M $4.050 / 1M
Total Input Tokens 69,747 86,746
Output Tokens 2,166 6,055
Reasoning Tokens 57,436 220,791
Response Time (avg) 8.20s 64.16s
Response Time (max) 76.68s 327.51s
Response Time (total) 180.47s 1411.59s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#12 Gemini 3.5 Flash

medium
Cost
$0.201
Time
112.9s
Tokens
22,371 tok

#32 Thinking Machines: Inkling

high
Cost
$0.005
Time
32.2s
Tokens
1,275 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.5 Flash 7.9 7.5 77.8% 1 12.63s 8,118 461 24,939
Inkling 8.5 10.0 66.7% 0 148.36s 7,374 430 78,982

Quick Compare

Switch Comparison Pair