- Rank
- #303
- Total Output Tokens
- 22,507
- Response Time (avg)
- 5.36s
- Total Cost
- $0.000
Dots 3 Note Preview vs GPT-4o-mini
The average score is effectively tied at 5.0 vs 5.0. Dots 3 Note Preview has the lower benchmark cost at $0.000 vs $0.024. GPT-4o-mini is faster at 4.16s vs 5.36s, with pass rates of 15.9% vs 21.7%.
Compared models
- Rank
- #304
- Total Output Tokens
- 3,627
- Response Time (avg)
- 4.16s
- Total Cost
- $0.024
Recommended model
GPT-4o-mini
It has the strongest score in this comparison (5.0) and the best overall balance of cost and response time across all 2 models.
Detailed comparison
| Metric | Dots 3 Note Preview Dots 3 Note Preview none Free Available | GPT-4o-mini GPT-4o-mini none |
|---|---|---|
| Score | 5.0 | 5.0 |
| Rank | #303 | #304 |
| Reliability | 10.0 | 10.0 |
| Consistency | 9.7 | 9.9 |
| Attempts | 69/69 | 69/69 |
| Tests Correct | ||
| Attempt pass rate | 15.9% | 21.7% |
| Flaky tests | 1 | 0 |
| Total Runs | 69 | 69 |
| Cost per result | 0.000 | 0.480 |
| Total Cost | $0.000 | $0.024 |
| Input Price | $0.000 / 1M | $0.150 / 1M |
| Output Price | $0.000 / 1M | $0.600 / 1M |
| Cache Read Price | N/A | $0.075 / 1M |
| Cache Write Price | N/A | N/A |
| Total Input Tokens | 225,907 | 145,180 |
| Output Tokens | 22,507 | 3,627 |
| Reasoning Tokens | 0 | 0 |
| Response Time (avg) | 5.36s | 4.16s |
| Response Time (max) | 45.71s | 39.99s |
| Response Time (total) | 123.38s | 70.70s |
| Parameters | 280B total (16B active) | ~8B |
| Availability | Open source | Closed |
Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#303 Dots 3 Note Preview
none- Cost
- $0.000
- Time
- 88.3s
- Tokens
- 10,531 tok
#304 GPT-4o-mini
none- Cost
- $0.001
- Time
- 6.6s
- Tokens
- 742 tok
Top Models by Score
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
| Agentic | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 3.9 | 9.6 | 0.0% | 0 | 45.71s | 109,752 | 2,828 | 0 | |
| GPT-4o-mini | 5.0 | 10.0 | 0.0% | 0 | 39.99s | 92,035 | 714 | 0 |
| Anti-AI Tricks | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 3.1 | 9.9 | 0.0% | 0 | 1.92s | 609 | 1,362 | 0 | |
| GPT-4o-mini | 4.8 | 10.0 | 25.0% | 0 | 1.34s | 618 | 186 | 0 |
| Coding | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 4.6 | 7.9 | 22.2% | 1 | 6.12s | 7,413 | 6,390 | 0 | |
| GPT-4o-mini | 3.2 | 9.6 | 0.0% | 0 | 1.63s | 7,314 | 367 | 0 |
| Combined | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 3.0 | 10.0 | 0.0% | 0 | 14.81s | 90,586 | 10,732 | 0 | |
| GPT-4o-mini | 3.0 | 10.0 | 0.0% | 0 | 6.32s | 29,916 | 1,497 | 0 |
| Data parsing and extraction | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 10.0 | 10.0 | 100.0% | 0 | 2.72s | 7,740 | 249 | 0 | |
| GPT-4o-mini | 10.0 | 10.0 | 100.0% | 0 | 1.27s | 7,161 | 183 | 0 |
| Domain specific | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 3.0 | 10.0 | 0.0% | 0 | 1.10s | 735 | 24 | 0 | |
| GPT-4o-mini | 3.0 | 10.0 | 0.0% | 0 | 743ms | 741 | 17 | 0 |
| General Intelligence | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 5.0 | 10.0 | 0.0% | 0 | 1.63s | 489 | 115 | 0 | |
| GPT-4o-mini | 4.0 | 10.0 | 0.0% | 0 | 909ms | 480 | 66 | 0 |
| Instructions following | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 4.8 | 10.0 | 0.0% | 0 | 1.06s | 666 | 72 | 0 | |
| GPT-4o-mini | 6.3 | 10.0 | 50.0% | 0 | 1.11s | 666 | 69 | 0 |
| Puzzle Solving | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 3.3 | 10.0 | 0.0% | 0 | 1.50s | 651 | 492 | 0 | |
| GPT-4o-mini | 3.5 | 10.0 | 0.0% | 0 | 1.21s | 651 | 308 | 0 |
| Tool Calling | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 10.0 | 10.0 | 100.0% | 0 | 4.04s | 7,059 | 231 | 0 | |
| GPT-4o-mini | 10.0 | 10.0 | 100.0% | 0 | 2.51s | 5,400 | 205 | 0 |
| Trivia | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Dots 3 Note Preview | 3.0 | 10.0 | 0.0% | 0 | 961ms | 207 | 12 | 0 | |
| GPT-4o-mini | 3.0 | 10.0 | 0.0% | 0 | 794ms | 198 | 15 | 0 |
Quick Compare
Switch Comparison Pair
KAT Coder AIR V2.5lowvsGPT-4o-mininoneDots 3 Note PreviewnoneFree AvailablevsMiniMax M2.7mediumDots 3 Note PreviewnoneFree AvailablevsKAT Coder AIR V2.5lowMiniMax M2.7mediumvsGPT-4o-mininoneDots 3 Note PreviewnoneFree AvailablevsMercury 2.5 PreviewlowMercury 2.5 PreviewlowvsGPT-4o-mininoneDots 3 Note PreviewnoneFree AvailablevsLaguna S 2.1mediumFree AvailableGPT-4o-mininonevsLaguna S 2.1highFree AvailableGPT-4o-mininonevsLaguna S 2.1mediumFree AvailableDots 3 Note PreviewnoneFree AvailablevsLaguna S 2.1highFree AvailableDots 3 Note PreviewnoneFree AvailablevsKAT Coder AIR V2.5mediumDots 3 Note PreviewnoneFree AvailablevsMercury 2.5low