Navigate
AI BENCHY
Advertise here

Granite 4.2 8B (medium) vs Inkling

The average score is effectively tied at 5.2 vs 5.2. Granite 4.2 8B (medium) has the lower benchmark cost at $0.009 vs $0.147. Inkling is faster at 3.47s vs 45.97s, with pass rates of 34.9% vs 28.8%.

Last updated at: 2026-09-02

Rank
#243
Total Output Tokens
19,330
Response Time (avg)
45.97s
Total Cost
$0.009
Rank
#242
Total Output Tokens
10,554
Response Time (avg)
3.47s
Total Cost
$0.147
Recommended model Granite 4.2 8B (medium)

It has the best score here (5.2), while costing about 18.0x less than Inkling.

Detailed comparison

Metric Granite 4.2 8B Granite 4.2 8B medium Release: 2026-09-02 Inkling Inkling none Release: 2026-07-18
Score 5.2 5.2
Rank #243 #242
Reliability 8.7 10.0
Consistency 9.6 9.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 34.9% 28.8%
Flaky tests 1 1
Total Runs 66 66
Cost per result 0.117 2.448
Total Cost $0.009 $0.147
Input Price $0.100 / 1M $1.000 / 1M
Output Price $0.150 / 1M $4.050 / 1M
Total Input Tokens 52,909 104,120
Output Tokens 7,058 10,554
Reasoning Tokens 12,272 0
Response Time (avg) 45.97s 3.47s
Response Time (max) 557.96s 48.02s
Response Time (total) 1011.42s 76.28s
Parameters 8B 975B total (41B active)
Availability Open source Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#243 IBM: Granite 4.2 8B

medium
Provider returned error
Cost
$0.000
Time
0.2s
Tokens
0 tok

#242 Thinking Machines: Inkling

none
Provider returned error
Cost
$0.000
Time
5.3s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Granite 4.2 8B 5.5 10.0 33.3% 0 12.09s 8,439 2,445 3,900
Inkling 4.5 10.0 0.0% 0 1.01s 7,356 436 0

Quick Compare

Switch Comparison Pair