Navigate
Advertise here

GPT-5.6 Luna vs Laguna XS 2.1

GPT-5.6 Luna leads on average score with 5.3 vs 5.3. Laguna XS 2.1 has the lower benchmark cost at $0.020 vs $0.042. GPT-5.6 Luna is faster at 3.78s vs 4.56s, with pass rates of 34.8% vs 29.0%.

Last updated at: 2026-10-07

Compared models

Rank
#297
Total Output Tokens
8,000
Response Time (avg)
3.78s
Total Cost
$0.042
Rank
#301
Total Output Tokens
23,173
Response Time (avg)
4.56s
Total Cost
$0.020
Recommended model GPT-5.6 Luna

It has the strongest score in this comparison (5.3) and the best overall balance of cost and response time across 2 models.

Detailed comparison

Metric GPT-5.6 Luna GPT-5.6 Luna none Release: 2026-07-09 Laguna XS 2.1 Laguna XS 2.1 none Release: 2026-07-02 Free Available
Score 5.3 5.3
Rank #297 #301
Reliability 10.0 10.0
Consistency 8.8 9.1
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 34.8% 29.0%
Flaky tests 3 3
Total Runs 69 69
Cost per result 2.814 0.380
Total Cost $0.042 $0.020
Input Price $0.200 / 1M $0.060 / 1M
Output Price $1.200 / 1M $0.120 / 1M
Cache Read Price $0.020 / 1M $0.030 / 1M
Cache Write Price $0.250 / 1M N/A
Total Input Tokens 230,744 270,072
Output Tokens 8,000 23,173
Reasoning Tokens 0 0
Response Time (avg) 3.78s 4.56s
Response Time (max) 53.77s 70.79s
Response Time (total) 86.83s 104.80s
Parameters ~400B total (~17B active) 33B total (3B active)
Availability Closed Weights available

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#297 GPT-5.6 Luna

none
Cost
$0.016
Time
15.8s
Tokens
2,685 tok

#301 Laguna XS 2.1

none
Cost
$0.001
Time
27.6s
Tokens
4,344 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
GPT-5.6 Luna 3.8 7.2 22.2% 1 980ms 7,302 459 0
Laguna XS 2.1 4.3 7.8 22.2% 1 623ms 7,995 562 0

Switch Comparison Pair