Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

IBM: Granite 4.1 8B vs Poolside: Laguna S 2.1

Laguna S 2.1 leads on average score with 4.5 vs 4.0. Granite 4.1 8B has the lower benchmark cost at $0.007 vs $0.025. Granite 4.1 8B is faster at 1.45s vs 11.67s, with pass rates of 9.1% vs 15.2%.

Recommended modelLaguna S 2.1It has the strongest score in this comparison (4.5) and the best overall balance of cost and response time across all 2 models.

Last updated at: 2026-07-22

Metric Granite 4.1 8B Granite 4.1 8B none Release: 2026-05-01 Laguna S 2.1 Laguna S 2.1 none Release: 2026-07-21 Free Available
Score 4.0 4.5
Rank #211 #200
Reliability 10.0 10.0
Consistency 10.0 8.8
Tests Correct
Attempt pass rate 9.1% 15.2%
Flaky tests 0 3
Total Runs 66 66
Cost per result 0.315 1.205
Total Cost $0.007 $0.025
Input Price $0.050 / 1M $0.100 / 1M
Output Price $0.100 / 1M $0.200 / 1M
Total Input Tokens 113,827 125,598
Output Tokens 5,996 57,634
Reasoning Tokens 0 0
Response Time (avg) 1.45s 11.67s
Response Time (max) 16.67s 200.52s
Response Time (total) 31.96s 256.80s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#211 IBM: Granite 4.1 8B

none
Cost
$0.001
Time
3.2s
Tokens
491 tok

#200 Laguna S 2.1

none
Cost
$0.001
Time
7.6s
Tokens
1,345 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Granite 4.1 8B 4.5 10.0 0.0% 0 775ms 8,344 525 0
Laguna S 2.1 3.4 9.7 0.0% 0 550ms 7,917 424 0

Quick Compare

Switch Comparison Pair