AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com
#390

GPT-6 Luna Decisions

OpenAI Release: 2026-10-07 Tested on: 2026-10-07 11:04 openai/gpt-6-luna-decisions::none
~400B total (~17B active)MoEClosedEstimated

Summary

GPT-6 Luna Decisions scores 2.4 on AI BENCHY and ranks #390. It has 10.0 reliability, a 13.0% pass rate, $0.005 total cost, and 495ms average response time.

What makes GPT-6 Luna Decisions unique: It is notably fast compared with similar models.

Model facts

Researched on 2026-10-07

Estimated
Parameters
~400B total (~17B active)
Architecture
MoE
Availability
Closed
License
-

Best estimate from public evidence; the vendor did not disclose every value. GPT-6 Luna served through the Decisions API, as identified by the exact OpenRouter checkpoint. OpenAI does not disclose parameter counts or architecture. Low-confidence estimates carry forward the existing GPT-6 Luna family record; these sources establish identity and closed API availability, not the estimated counts. The Decisions route returns typed choices and probabilities, without text generation or exposed reasoning controls.

Score

2.4

Consistency

5.2

Total Output Tokens

0

Total Input Tokens

47,583

Input Price

$0.100 / 1M

Output Price

$0.000 / 1M

Cache Read Price

$0.000 / 1M

Cache Write Price

$0.000 / 1M

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Tests Correct

Wrong Tests: 9

Attempt pass rate: 13.0%

Benchmark coverage: 36/69 attempts. Choices supplied by adapters. Supports 12/23 tests. Other tests are unsupported, not failed. Score uses the full suite.

Flaky tests

0

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

495ms

Response Time (max): 842ms

Response Time (total): 5.94s

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Agentic 0.0 0.0
Anti-AI Tricks 2.3 7.5
Coding 2.5 6.7
Combined 0.0 0.0
Data parsing and extraction 10.0 10.0
Domain specific 5.3 10.0
General Intelligence 0.0 0.0
Instructions following 1.5 5.0
Puzzle Solving 1.7 3.3
Tool Calling 0.0 0.0
Trivia 0.0 0.0

Compared models