- Rank
- #368
- Total Output Tokens
- 939,018
- Response Time (avg)
- 72.24s
- Total Cost
- $0.000
Ling 3.0 Tiny (high) vs GPT-6 Luna Decisions (default)
Ling 3.0 Tiny (high) leads on average score with 3.7 vs 2.4. Ling 3.0 Tiny (high) has the lower benchmark cost at $0.000 vs $0.005. GPT-6 Luna Decisions (default) is faster at 495ms vs 72.24s, with pass rates of 23.2% vs 13.0%.
Compared models
- Rank
- #390
- Total Output Tokens
- 0
- Response Time (avg)
- 495ms
- Total Cost
- $0.005
Recommended model
Ling 3.0 Tiny (high)
It has the strongest score in this comparison (3.7) and the best overall balance of cost and response time across all 2 models.
Detailed comparison
| Metric | Ling 3.0 Tiny Ling 3.0 Tiny high | GPT-6 Luna Decisions GPT-6 Luna Decisions default |
|---|---|---|
| Score | 3.7 | 2.4 |
| Rank | #368 | #390 |
| Reliability | 10.0 | 10.0 |
| Consistency | 8.9 | 5.2 |
| Attempts | 69/69 | 36/69 |
| Tests Correct | ||
| Attempt pass rate | 23.2% | 13.0% |
| Flaky tests | 3 | 0 |
| Total Runs | 69 | 36 |
| Cost per result | 0.000 | 0.159 |
| Total Cost | $0.000 | $0.005 |
| Input Price | $0.000 / 1M | $0.100 / 1M |
| Output Price | $0.000 / 1M | $0.000 / 1M |
| Cache Read Price | N/A | $0.000 / 1M |
| Cache Write Price | N/A | $0.000 / 1M |
| Total Input Tokens | 115,314 | 47,583 |
| Output Tokens | 133,965 | 0 |
| Reasoning Tokens | 816,718 | 0 |
| Response Time (avg) | 72.24s | 495ms |
| Response Time (max) | 345.15s | 842ms |
| Response Time (total) | 1589.23s | 5.94s |
| Parameters | 7.9B total (1.3B active) | ~400B total (~17B active) |
| Availability | Open source | Closed |
Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#368 Ling 3.0 Tiny
high
No output was saved. The original provider response or failure reason is unavailable.
- Cost
- $0.000
- Time
- 169.5s
- Tokens
- 5,252 tok
#390 GPT-6 Luna Decisions
default
No showcase result has been generated for this model yet.
- Cost
- N/A
- Time
- -
- Tokens
- 0 tok
Top Models by Score
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
| Agentic | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 28ms | 0 | 0 | 0 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Anti-AI Tricks | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 6.4 | 7.9 | 58.3% | 1 | 24.93s | 750 | 5,102 | 94,893 | |
| GPT-6 Luna Decisions | 2.3 | 7.5 | 0.0% | 0 | 436ms | 6,279 | 0 | 0 |
| Coding | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 2.9 | 10.0 | 0.0% | 0 | 177.65s | 7,767 | 59,320 | 177,922 | |
| GPT-6 Luna Decisions | 2.5 | 6.7 | 0.0% | 0 | 447ms | 22,197 | 0 | 0 |
| Combined | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 179.85s | 87,658 | 19,652 | 152,168 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Data parsing and extraction | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 2.8 | 10.0 | 0.0% | 0 | 4.16s | 8,016 | 1,561 | 5,372 | |
| GPT-6 Luna Decisions | 10.0 | 10.0 | 100.0% | 0 | 411ms | 10,998 | 0 | 0 |
| Domain specific | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 2.9 | 7.2 | 11.1% | 1 | 70.12s | 648 | 22,445 | 155,528 | |
| GPT-6 Luna Decisions | 5.3 | 10.0 | 33.3% | 0 | 554ms | 3,063 | 0 | 0 |
| General Intelligence | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 5.4 | 2.5 | 66.7% | 1 | 60.50s | 546 | 4,900 | 61,589 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Instructions following | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 6.8 | 9.4 | 50.0% | 0 | 3.10s | 744 | 618 | 1,444 | |
| GPT-6 Luna Decisions | 1.5 | 5.0 | 0.0% | 0 | 842ms | 4,440 | 0 | 0 |
| Puzzle Solving | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 54.15s | 731 | 10,323 | 78,452 | |
| GPT-6 Luna Decisions | 1.7 | 3.3 | 0.0% | 0 | 414ms | 606 | 0 | 0 |
| Tool Calling | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 10.0 | 10.0 | 100.0% | 0 | 5.76s | 8,226 | 379 | 711 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Trivia | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 213.35s | 228 | 9,665 | 88,639 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
Quick Compare
Switch Comparison Pair
Ling 3.0 TinyhighvsNemotron 3.5 LightningnoneFree AvailableGranite 4.1 8bdefaultvsLing 3.0 TinyhighLing 3.0 TinyhighvsGrok 4.20noneCommand A+mediumvsLing 3.0 TinyhighLing 3.0 TinyhighvsQwen3.5-9BmediumMercury 2.5nonevsLing 3.0 TinyhighLing 3.0 Tinyhighvsgpt-oss-120bnoneLing 3.0 TinyhighvsGLM 4.7 FlashmediumCommand A+lowvsGPT-6 Luna DecisionsdefaultCommand A+nonevsGPT-6 Luna DecisionsdefaultLing 3.0 TinyhighvsMiniMax M2.5mediumGemini 3 Flash PreviewhighvsGPT-6 Luna Decisionsdefault