Navigate
AI BENCHY
Advertise here

Muse Spark 1.1 (low) vs Inkling Small (high)

Inkling Small (high) leads on average score with 8.4 vs 8.3. Inkling Small (high) has the lower benchmark cost at $0.416 vs $0.647. Muse Spark 1.1 (low) is faster at 11.45s vs 31.70s, with pass rates of 69.7% vs 69.7%.

Last updated at: 2026-08-01

Rank
#33
Total Output Tokens
110,314
Response Time (avg)
11.45s
Total Cost
$0.647
Rank
#30
Total Output Tokens
310,388
Response Time (avg)
31.70s
Total Cost
$0.416
Recommended model Muse Spark 1.1 (low)

Its score stays close to the best score here (8.3 vs 8.4), while responding about 2.8x faster than Inkling Small (high).

Detailed comparison

Metric Muse Spark 1.1 Muse Spark 1.1 low Release: 2026-07-16 Inkling Small Inkling Small high Release: 2026-08-01
Score 8.3 8.4
Rank #33 #30
Reliability 10.0 10.0
Consistency 8.5 9.2
Tests Correct
Attempt pass rate 69.7% 69.7%
Flaky tests 4 2
Total Runs 66 66
Cost per result 4.975 2.967
Total Cost $0.647 $0.416
Input Price $1.250 / 1M $0.500 / 1M
Output Price $4.250 / 1M $1.200 / 1M
Total Input Tokens 142,298 85,660
Output Tokens 10,847 6,092
Reasoning Tokens 99,467 304,296
Response Time (avg) 11.45s 31.70s
Response Time (max) 54.15s 122.60s
Response Time (total) 251.92s 697.30s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#33 Muse Spark 1.1

low
Cost
$0.020
Time
39.2s
Tokens
4,947 tok

#30 Thinking Machines: Inkling Small

high
Cost
$0.006
Time
34.7s
Tokens
4,735 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Muse Spark 1.1 10.0 10.0 100.0% 0 10.34s 8,562 394 14,503
Inkling Small 10.0 10.0 100.0% 0 93.23s 7,374 463 120,474

Quick Compare

Switch Comparison Pair