Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

IBM: Granite 4.1 8B vs NVIDIA: Nemotron 3 Super

Last updated at: 2026-05-29

Metric Granite 4.1 8B Granite 4.1 8B none Release: 2026-05-01 Nemotron 3 Super Nemotron 3 Super none Release: 2026-03-11 Free Available
Score 4.1 5.0
Rank #158 #139
Reliability 10.0 10.0
Consistency 10.0 8.7
Tests Correct
Attempt pass rate 10.0% 33.3%
Flaky tests 0 3
Total Runs 60 60
Cost per result 0.122 0.034
Total Cost $0.003 $0.002
Input Price $0.050 / 1M $0.090 / 1M
Output Price $0.100 / 1M $0.450 / 1M
Output Tokens 2,743 6,186
Reasoning Tokens 0 0
Response Time (avg) 719ms 5.47s
Response Time (max) 2.17s 16.45s
Response Time (total) 14.37s 109.43s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 4.9 10.0 25.0% 0 844ms 903 0
Nemotron 3 Super 4.8 10.0 25.0% 0 4.46s 2,322 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 5.2 10.0 0.0% 0 706ms 357 0
Nemotron 3 Super 3.4 5.8 16.7% 1 3.02s 562 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 3.0 10.0 0.0% 0 1.88s 396 0
Nemotron 3 Super 3.0 10.0 0.0% 0 16.45s 617 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 3.0 10.0 0.0% 0 575ms 195 0
Nemotron 3 Super 10.0 10.0 100.0% 0 7.92s 249 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 3.0 10.0 0.0% 0 357ms 24 0
Nemotron 3 Super 3.6 7.2 22.2% 1 6.23s 26 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 4.0 10.0 0.0% 0 499ms 115 0
Nemotron 3 Super 4.6 10.0 0.0% 0 950ms 134 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 3.6 9.9 0.0% 0 344ms 66 0
Nemotron 3 Super 6.3 10.0 50.0% 0 804ms 66 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 3.2 10.0 0.0% 0 608ms 432 0
Nemotron 3 Super 5.5 10.0 33.3% 0 2.36s 1,125 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 10.0 10.0 100.0% 0 2.17s 243 0
Nemotron 3 Super 4.7 1.6 66.7% 1 16.00s 281 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Granite 4.1 8B 3.0 10.0 0.0% 0 306ms 12 0
Nemotron 3 Super 3.0 10.0 0.0% 0 8.94s 804 0

Quick Compare

Switch Comparison Pair