- Rank
- #370
- Total Output Tokens
- 911,748
- Response Time (avg)
- 64.99s
- Total Cost
- $0.000
Ling 3.0 Tiny (medium) vs GPT-6 Luna Decisions (default)
Ling 3.0 Tiny (medium) leads on average score with 3.6 vs 2.4. Ling 3.0 Tiny (medium) has the lower benchmark cost at $0.000 vs $0.005. GPT-6 Luna Decisions (default) is faster at 495ms vs 64.99s, with pass rates of 23.2% vs 13.0%.
Compared models
- Rank
- #390
- Total Output Tokens
- 0
- Response Time (avg)
- 495ms
- Total Cost
- $0.005
Recommended model
Ling 3.0 Tiny (medium)
It has the strongest score in this comparison (3.6) and the best overall balance of cost and response time across all 2 models.
Detailed comparison
| Metric | Ling 3.0 Tiny Ling 3.0 Tiny medium | GPT-6 Luna Decisions GPT-6 Luna Decisions default |
|---|---|---|
| Score | 3.6 | 2.4 |
| Rank | #370 | #390 |
| Reliability | 9.6 | 10.0 |
| Consistency | 8.9 | 5.2 |
| Attempts | 69/69 | 36/69 |
| Tests Correct | ||
| Attempt pass rate | 23.2% | 13.0% |
| Flaky tests | 3 | 0 |
| Total Runs | 69 | 36 |
| Cost per result | 0.000 | 0.159 |
| Total Cost | $0.000 | $0.005 |
| Input Price | $0.000 / 1M | $0.100 / 1M |
| Output Price | $0.000 / 1M | $0.000 / 1M |
| Cache Read Price | N/A | $0.000 / 1M |
| Cache Write Price | N/A | $0.000 / 1M |
| Total Input Tokens | 103,898 | 47,583 |
| Output Tokens | 148,304 | 0 |
| Reasoning Tokens | 774,972 | 0 |
| Response Time (avg) | 64.99s | 495ms |
| Response Time (max) | 262.20s | 842ms |
| Response Time (total) | 1429.68s | 5.94s |
| Parameters | 7.9B total (1.3B active) | ~400B total (~17B active) |
| Availability | Open source | Closed |
Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#370 Ling 3.0 Tiny
medium
No output was saved. The original provider response or failure reason is unavailable.
- Cost
- $0.000
- Time
- 177.4s
- Tokens
- 6,873 tok
#390 GPT-6 Luna Decisions
default
No showcase result has been generated for this model yet.
- Cost
- N/A
- Time
- -
- Tokens
- 0 tok
Top Models by Score
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
| Agentic | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 24ms | 0 | 0 | 0 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Anti-AI Tricks | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 6.5 | 10.0 | 50.0% | 0 | 23.73s | 750 | 5,103 | 94,884 | |
| GPT-6 Luna Decisions | 2.3 | 7.5 | 0.0% | 0 | 436ms | 6,279 | 0 | 0 |
| Coding | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 2.9 | 10.0 | 0.0% | 0 | 211.99s | 7,948 | 70,838 | 206,006 | |
| GPT-6 Luna Decisions | 2.5 | 6.7 | 0.0% | 0 | 447ms | 22,197 | 0 | 0 |
| Combined | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 139.78s | 76,108 | 17,871 | 126,497 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Data parsing and extraction | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 2.7 | 5.7 | 16.7% | 1 | 4.41s | 8,016 | 2,014 | 6,755 | |
| GPT-6 Luna Decisions | 10.0 | 10.0 | 100.0% | 0 | 411ms | 10,998 | 0 | 0 |
| Domain specific | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 81.40s | 637 | 21,709 | 187,100 | |
| GPT-6 Luna Decisions | 5.3 | 10.0 | 33.3% | 0 | 554ms | 3,063 | 0 | 0 |
| General Intelligence | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.8 | 2.5 | 33.3% | 1 | 31.77s | 546 | 4,275 | 30,975 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Instructions following | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 6.8 | 10.0 | 50.0% | 0 | 3.00s | 744 | 610 | 1,436 | |
| GPT-6 Luna Decisions | 1.5 | 5.0 | 0.0% | 0 | 842ms | 4,440 | 0 | 0 |
| Puzzle Solving | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.6 | 7.2 | 22.2% | 1 | 34.65s | 747 | 15,561 | 85,029 | |
| GPT-6 Luna Decisions | 1.7 | 3.3 | 0.0% | 0 | 414ms | 606 | 0 | 0 |
| Tool Calling | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 10.0 | 10.0 | 100.0% | 0 | 4.91s | 8,226 | 380 | 708 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
| Trivia | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Ling 3.0 Tiny | 3.0 | 10.0 | 0.0% | 0 | 100.96s | 176 | 9,943 | 35,582 | |
| GPT-6 Luna Decisions | 0.0 | 0.0 | 0.0% | 0 | 0ms | 0 | 0 | 0 |
Quick Compare
Switch Comparison Pair
Ling 3.0 TinymediumvsNemotron 3.5 LightningnoneFree AvailableGranite 4.1 8bdefaultvsLing 3.0 TinymediumLing 3.0 TinymediumvsGrok 4.20noneMercury 2.5nonevsLing 3.0 TinymediumCommand A+highvsLing 3.0 TinymediumLing 3.0 Tinymediumvsgpt-oss-120bnoneCommand A+lowvsGPT-6 Luna DecisionsdefaultCommand A+nonevsGPT-6 Luna DecisionsdefaultGemini 3 Flash PreviewhighvsGPT-6 Luna DecisionsdefaultCommand A+nonevsLing 3.0 TinymediumCommand A+lowvsLing 3.0 TinymediumMercury 2nonevsLing 3.0 Tinymedium