Navigate
Advertise here

GPT-5.4 Nano (medium) vs GPT-5.6 Luna (high)

The average score is effectively tied at 8.0 vs 8.0. GPT-5.6 Luna (high) has the lower benchmark cost at $0.118 vs $0.235. GPT-5.4 Nano (medium) is faster at 15.95s vs 18.39s, with pass rates of 66.7% vs 71.0%.

Last updated at: 2026-10-07

Compared models

Rank
#98
Total Output Tokens
110,910
Response Time (avg)
15.95s
Total Cost
$0.235
Rank
#96
Total Output Tokens
139,395
Response Time (avg)
18.39s
Total Cost
$0.118
Recommended model GPT-5.6 Luna (high)

It has the best score here (8.0), while costing about 2.0x less than GPT-5.4 Nano (medium).

Detailed comparison

Metric GPT-5.4 Nano GPT-5.4 Nano medium Release: 2026-03-17 GPT-5.6 Luna GPT-5.6 Luna high Release: 2026-07-09
Score 8.0 8.0
Rank #98 #96
Reliability 10.0 10.0
Consistency 8.6 8.6
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 66.7% 71.0%
Flaky tests 4 4
Total Runs 69 69
Cost per result 1.432 5.882
Total Cost $0.235 $0.118
Input Price $0.200 / 1M $0.200 / 1M
Output Price $1.250 / 1M $1.200 / 1M
Cache Read Price $0.020 / 1M $0.020 / 1M
Cache Write Price N/A $0.250 / 1M
Total Input Tokens 237,088 197,896
Output Tokens 8,235 6,211
Reasoning Tokens 102,675 133,184
Response Time (avg) 15.95s 18.39s
Response Time (max) 94.06s 111.09s
Response Time (total) 366.93s 422.93s
Parameters ~120B total (~5B active) ~400B total (~17B active)
Availability Closed Closed

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#98 GPT-5.4 Nano

medium
Cost
$0.007
Time
24.6s
Tokens
4,943 tok

#96 GPT-5.6 Luna

high
Cost
$0.033
Time
34.2s
Tokens
5,484 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
GPT-5.4 Nano 6.1 4.7 66.7% 2 19.12s 7,305 516 20,778
GPT-5.6 Luna 5.5 4.7 55.6% 2 15.63s 7,302 510 17,746

Switch Comparison Pair