Navigate
AI BENCHY
Advertise here

Granite 4.2 8B vs Laguna S 2.1

Laguna S 2.1 leads on average score with 4.5 vs 4.3. Laguna S 2.1 has the lower benchmark cost at $0.022 vs $0.026. Laguna S 2.1 is faster at 11.67s vs 86.27s, with pass rates of 16.7% vs 15.2%.

Last updated at: 2026-09-02

Rank
#285
Total Output Tokens
130,805
Response Time (avg)
86.27s
Total Cost
$0.026
Rank
#279
Total Output Tokens
57,634
Response Time (avg)
11.67s
Total Cost
$0.022
Recommended model Laguna S 2.1

It has the best score here (4.5), while responding about 7.4x faster than Granite 4.2 8B.

Detailed comparison

Metric Granite 4.2 8B Granite 4.2 8B none Release: 2026-09-02 Laguna S 2.1 Laguna S 2.1 none Release: 2026-07-21 Free Available
Score 4.3 4.5
Rank #285 #279
Reliability 8.6 10.0
Consistency 9.3 8.8
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 16.7% 15.2%
Flaky tests 2 3
Total Runs 66 66
Cost per result 0.847 1.205
Total Cost $0.026 $0.022
Input Price $0.100 / 1M $0.090 / 1M
Output Price $0.150 / 1M $0.180 / 1M
Total Input Tokens 57,786 125,604
Output Tokens 130,805 57,634
Reasoning Tokens 0 0
Response Time (avg) 86.27s 11.67s
Response Time (max) 747.17s 200.52s
Response Time (total) 1897.96s 256.71s
Parameters 8B 118B total (8B active)
Availability Open source Weights available

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#285 IBM: Granite 4.2 8B

none
Provider returned error
Cost
$0.000
Time
0.2s
Tokens
0 tok

#279 Laguna S 2.1

none
Cost
$0.001
Time
7.6s
Tokens
1,345 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Granite 4.2 8B 5.5 10.0 33.3% 0 105.60s 8,358 47,944 0
Laguna S 2.1 3.4 9.7 0.0% 0 550ms 7,917 424 0

Quick Compare

Switch Comparison Pair