- Rank
- #171
- Total Output Tokens
- 150,107
- Response Time (avg)
- 6.66s
- Total Cost
- $0.037
Mercury 2.5 Preview (medium) vs Space Bunny Alpha (max)
Mercury 2.5 Preview (medium) leads on average score with 6.8 vs 6.7. Space Bunny Alpha (max) has the lower benchmark cost at $0.000 vs $0.037. Mercury 2.5 Preview (medium) is faster at 6.66s vs 36.24s, with pass rates of 66.7% vs 59.4%.
Compared models
- Rank
- #173
- Total Output Tokens
- 219,014
- Response Time (avg)
- 36.24s
- Total Cost
- $0.000
Recommended model
Mercury 2.5 Preview (medium)
It has the best score here (6.8), while responding about 5.4x faster than Space Bunny Alpha (max).
Detailed comparison
| Metric | Mercury 2.5 Preview Mercury 2.5 Preview medium | Space Bunny Alpha Space Bunny Alpha max |
|---|---|---|
| Score | 6.8 | 6.7 |
| Rank | #171 | #173 |
| Reliability | 10.0 | 10.0 |
| Consistency | 9.6 | 8.3 |
| Attempts | 69/69 | 69/69 |
| Tests Correct | ||
| Attempt pass rate | 66.7% | 59.4% |
| Flaky tests | 1 | 5 |
| Total Runs | 69 | 69 |
| Cost per result | 0.244 | 0.000 |
| Total Cost | $0.037 | $0.000 |
| Input Price | $0.000 / 1M | $0.000 / 1M |
| Output Price | $0.000 / 1M | $0.000 / 1M |
| Total Input Tokens | 397,291 | 315,136 |
| Output Tokens | 6,355 | 219,014 |
| Reasoning Tokens | 143,752 | 0 |
| Response Time (avg) | 6.66s | 36.24s |
| Response Time (max) | 74.63s | 320.00s |
| Response Time (total) | 153.29s | 833.43s |
| Parameters | ~100B | - |
| Availability | Closed | Closed |
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#171 Mercury 2.5 Preview
medium- Cost
- $0.001
- Time
- 2.7s
- Tokens
- 2,241 tok
#173 Space Bunny Alpha
max- Cost
- $0.000
- Time
- 183.9s
- Tokens
- 19,570 tok
Top Models by Score
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
| Agentic | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 3.9 | 9.6 | 0.0% | 0 | 74.63s | 284,578 | 919 | 22,154 | |
| Space Bunny Alpha | 6.1 | 3.1 | 66.7% | 1 | 90.59s | 182,282 | 11,433 | 0 |
| Anti-AI Tricks | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 10.0 | 10.0 | 100.0% | 0 | 1.43s | 781 | 203 | 7,921 | |
| Space Bunny Alpha | 6.4 | 7.9 | 58.3% | 1 | 2.61s | 2,130 | 2,397 | 0 |
| Coding | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 7.8 | 10.0 | 66.7% | 0 | 3.86s | 7,839 | 552 | 22,126 | |
| Space Bunny Alpha | 5.7 | 7.1 | 44.4% | 1 | 17.08s | 8,493 | 16,335 | 0 |
| Combined | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 6.5 | 10.0 | 50.0% | 0 | 13.86s | 92,663 | 3,607 | 41,588 | |
| Space Bunny Alpha | 6.5 | 10.0 | 50.0% | 0 | 61.05s | 98,873 | 41,023 | 0 |
| Data parsing and extraction | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 10.0 | 10.0 | 100.0% | 0 | 2.77s | 8,292 | 514 | 11,354 | |
| Space Bunny Alpha | 10.0 | 10.0 | 100.0% | 0 | 2.64s | 7,890 | 1,077 | 0 |
| Domain specific | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 5.3 | 7.2 | 44.4% | 1 | 3.72s | 853 | 62 | 17,120 | |
| Space Bunny Alpha | 7.7 | 10.0 | 66.7% | 0 | 140.04s | 1,869 | 114,458 | 0 |
| General Intelligence | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 5.0 | 10.0 | 0.0% | 0 | 2.09s | 538 | 99 | 3,375 | |
| Space Bunny Alpha | 5.3 | 10.0 | 0.0% | 0 | 8.15s | 855 | 2,091 | 0 |
| Instructions following | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 9.8 | 10.0 | 100.0% | 0 | 1.35s | 744 | 96 | 3,987 | |
| Space Bunny Alpha | 5.9 | 2.6 | 66.7% | 2 | 3.78s | 1,425 | 1,845 | 0 |
| Puzzle Solving | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 10.0 | 10.0 | 100.0% | 0 | 1.87s | 771 | 274 | 7,323 | |
| Space Bunny Alpha | 7.8 | 9.9 | 66.7% | 0 | 9.15s | 1,782 | 7,206 | 0 |
| Tool Calling | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 3.0 | 10.0 | 0.0% | 0 | 2.37s | 0 | 0 | 0 | |
| Space Bunny Alpha | 10.0 | 10.0 | 100.0% | 0 | 5.58s | 8,961 | 1,076 | 0 |
| Trivia | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Mercury 2.5 Preview | 3.0 | 10.0 | 0.0% | 0 | 4.13s | 232 | 29 | 6,804 | |
| Space Bunny Alpha | 3.0 | 10.0 | 0.0% | 0 | 84.94s | 576 | 20,073 | 0 |
Quick Compare
Switch Comparison Pair
Mercury 2.5 PreviewmediumvsInklinglowFree AvailableMercury 2.5 PreviewmediumvsGPT-6 SolnoneSpace Bunny AlphamaxvsSolar Pro 4xhighSpace Bunny AlphamaxvsGrok 4.7xhighSpace Bunny AlphamaxvsGrok Build 0.1mediumClaude Opus 4.8nonevsMercury 2.5 PreviewmediumClaude Sonnet 5.5lowvsMercury 2.5 PreviewmediumMercury 2.5 PreviewmediumvsGPT-5.6 LunalowGPT-6 SolnonevsSpace Bunny AlphamaxQwen3.5-FlashnonevsSpace Bunny AlphamaxEmber-1highvsMercury 2.5 PreviewmediumSpace Bunny AlphamaxvsSolar Pro 4low