Navigate
AI BENCHY
Advertise here

Gemma 4 26B A4B (medium) vs Inkling Small (medium)

The average score is effectively tied at 6.6 vs 6.6. Gemma 4 26B A4B (medium) has the lower benchmark cost at $0.089 vs $0.117. Inkling Small (medium) is faster at 6.27s vs 103.83s, with pass rates of 66.7% vs 56.1%.

Last updated at: 2026-08-01

Rank
#113
Total Output Tokens
247,527
Response Time (avg)
103.83s
Total Cost
$0.089
Rank
#112
Total Output Tokens
55,321
Response Time (avg)
6.27s
Total Cost
$0.117
Recommended model Inkling Small (medium)

It has the best score here (6.6), while responding about 16.6x faster than Gemma 4 26B A4B (medium).

Detailed comparison

Metric Gemma 4 26B A4B Gemma 4 26B A4B medium Release: 2026-04-03 Free Available Inkling Small Inkling Small medium Release: 2026-08-01
Score 6.6 6.6
Rank #113 #112
Reliability 9.4 10.0
Consistency 9.2 8.6
Tests Correct
Attempt pass rate 66.7% 56.1%
Flaky tests 2 4
Total Runs 66 66
Cost per result 0.643 1.166
Total Cost $0.089 $0.117
Input Price $0.070 / 1M $0.500 / 1M
Output Price $0.340 / 1M $1.200 / 1M
Total Input Tokens 77,550 100,246
Output Tokens 28,036 6,422
Reasoning Tokens 219,491 48,899
Response Time (avg) 103.83s 6.27s
Response Time (max) 912.19s 17.26s
Response Time (total) 2180.47s 138.02s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#113 Gemma 4 26B A4B

medium
Invalid SVG
Cost
$0.000
Time
300.0s
Tokens
0 tok

#112 Thinking Machines: Inkling Small

medium
Cost
$0.003
Time
13.4s
Tokens
1,753 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemma 4 26B A4B 2.9 10.0 0.0% 0 272.54s 5,062 14,838 44,567
Inkling Small 7.8 10.0 66.7% 0 10.49s 7,374 434 13,783

Quick Compare

Switch Comparison Pair