Navigate
Advertise here

Ling 3.0 Tiny (low) vs GPT-6 Luna Decisions (default)

Ling 3.0 Tiny (low) leads on average score with 3.7 vs 2.4. Ling 3.0 Tiny (low) has the lower benchmark cost at $0.000 vs $0.005. GPT-6 Luna Decisions (default) is faster at 495ms vs 63.15s, with pass rates of 18.8% vs 13.0%.

Last updated at: 2026-10-07

Compared models

Rank
#369
Total Output Tokens
868,004
Response Time (avg)
63.15s
Total Cost
$0.000
Rank
#390
Total Output Tokens
0
Response Time (avg)
495ms
Total Cost
$0.005
Recommended model Ling 3.0 Tiny (low)

It has the strongest score in this comparison (3.7) and the best overall balance of cost and response time across all 2 models.

Detailed comparison

Metric Ling 3.0 Tiny Ling 3.0 Tiny low Release: 2026-08-07 GPT-6 Luna Decisions GPT-6 Luna Decisions default Release: 2026-10-07
Score 3.7 2.4
Rank #369 #390
Reliability 9.6 10.0
Consistency 9.7 5.2
Attempts 69/69 36/69
Tests Correct
Attempt pass rate 18.8% 13.0%
Flaky tests 1 0
Total Runs 69 36
Cost per result 0.000 0.159
Total Cost $0.000 $0.005
Input Price $0.000 / 1M $0.100 / 1M
Output Price $0.000 / 1M $0.000 / 1M
Cache Read Price N/A $0.000 / 1M
Cache Write Price N/A $0.000 / 1M
Total Input Tokens 99,662 47,583
Output Tokens 146,555 0
Reasoning Tokens 732,977 0
Response Time (avg) 63.15s 495ms
Response Time (max) 252.08s 842ms
Response Time (total) 1389.35s 5.94s
Parameters 7.9B total (1.3B active) ~400B total (~17B active)
Availability Open source Closed

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#369 Ling 3.0 Tiny

low
No output was saved. The original provider response or failure reason is unavailable.
Cost
$0.000
Time
200.0s
Tokens
5,030 tok

#390 GPT-6 Luna Decisions

default
No showcase result has been generated for this model yet.
Cost
N/A
Time
-
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling 3.0 Tiny 2.9 10.0 0.0% 0 170.73s 7,619 56,012 170,220
GPT-6 Luna Decisions 2.5 6.7 0.0% 0 447ms 22,197 0 0

Quick Compare

Switch Comparison Pair