Navigate
Advertise here

Claude Opus 5 vs Ember-1 (low)

Claude Opus 5 leads on average score with 7.5 vs 7.5. Ember-1 (low) has the lower benchmark cost at $0.662 vs $1.550. Claude Opus 5 is faster at 7.65s vs 30.98s, with pass rates of 57.6% vs 66.7%.

Last updated at: 2026-09-28

Compared models

Rank
#115
Total Output Tokens
24,866
Response Time (avg)
7.65s
Total Cost
$1.550
Rank
#120
Total Output Tokens
37,467
Response Time (avg)
30.98s
Total Cost
$0.662
Recommended model Claude Opus 5

It has the best score here (7.5), while responding about 4.0x faster than Ember-1 (low).

Detailed comparison

Metric Claude Opus 5 Claude Opus 5 none Release: 2026-07-25 Ember-1 Ember-1 low Release: 2026-09-28
Score 7.5 7.5
Rank #115 #120
Reliability 10.0 8.1
Consistency 9.6 8.5
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 57.6% 66.7%
Flaky tests 1 4
Total Runs 66 66
Cost per result 12.912 5.085
Total Cost $1.550 $0.662
Input Price $5.000 / 1M $3.000 / 1M
Output Price $25.000 / 1M $15.000 / 1M
Total Input Tokens 185,547 33,011
Output Tokens 24,866 4,780
Reasoning Tokens 0 32,687
Response Time (avg) 7.65s 30.98s
Response Time (max) 47.72s 93.10s
Response Time (total) 168.35s 681.50s
Parameters ~5T total (~500B active) ~2.8T total (~104B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#115 Claude Opus 5

none
Cost
$0.271
Time
149.3s
Tokens
10,962 tok

#120 Ember-1

low
Cost
$0.016
Time
24.2s
Tokens
1,165 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Opus 5 5.9 9.7 33.3% 0 5.54s 10,590 2,886 0
Ember-1 10.0 10.0 100.0% 0 37.91s 8,097 2,518 13,142

Quick Compare

Switch Comparison Pair