Navigate
Advertise here

Command A+ (high) vs Granite 4.1 8B

Granite 4.1 8B leads on average score with 4.0 vs 3.0. Command A+ (high) has the lower benchmark cost at $0.000 vs $0.007. Granite 4.1 8B is faster at 1.44s vs 86.95s, with pass rates of 0.0% vs 9.1%.

Last updated at: 2026-09-23

Compared models

Rank
#353
Total Output Tokens
0
Response Time (avg)
86.95s
Total Cost
$0.000
Rank
#339
Total Output Tokens
5,996
Response Time (avg)
1.44s
Total Cost
$0.007
Recommended model Granite 4.1 8B

It has the best score here (4.0), while responding about 60.3x faster than Command A+ (high).

Detailed comparison

Metric Command A+ Command A+ high Release: 2026-09-23 Granite 4.1 8B Granite 4.1 8B none Release: 2026-05-01
Score 3.0 4.0
Rank #353 #339
Reliability 0.0 10.0
Consistency 10.0 10.0
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 0.0% 9.1%
Flaky tests 0 0
Total Runs 66 66
Cost per result 0.000 0.315
Total Cost $0.000 $0.007
Input Price $0.300 / 1M $0.050 / 1M
Output Price $1.500 / 1M $0.100 / 1M
Total Input Tokens 0 113,836
Output Tokens 0 5,996
Reasoning Tokens 0 0
Response Time (avg) 86.95s 1.44s
Response Time (max) 114.87s 16.67s
Response Time (total) 1913.01s 31.75s
Parameters 218B total (25B active) 8B
Availability Open source Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#353 Command A+

high
Provider returned error
Cost
$0.000
Time
0.2s
Tokens
0 tok

#339 IBM: Granite 4.1 8B

none
Cost
$0.001
Time
3.2s
Tokens
491 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Command A+ 3.0 10.0 0.0% 0 92.74s 0 0 0
Granite 4.1 8B 4.5 10.0 0.0% 0 775ms 8,344 525 0

Quick Compare

Switch Comparison Pair