Navigate
Advertise here

GPT-5.6 Luna vs MiMo-V2.6-Flash

MiMo-V2.6-Flash leads on average score with 5.4 vs 5.4. GPT-5.6 Luna has the lower benchmark cost at $0.029 vs $0.037. GPT-5.6 Luna is faster at 1.50s vs 3.64s, with pass rates of 36.4% vs 31.8%.

Last updated at: 2026-09-21

Compared models

Rank
#263
Total Output Tokens
6,709
Response Time (avg)
1.50s
Total Cost
$0.029
Rank
#258
Total Output Tokens
74,105
Response Time (avg)
3.64s
Total Cost
$0.037
Recommended model GPT-5.6 Luna

Its score stays close to the best score here (5.4 vs 5.4), while responding about 2.4x faster than MiMo-V2.6-Flash.

Detailed comparison

Metric GPT-5.6 Luna GPT-5.6 Luna none Release: 2026-07-09 MiMo-V2.6-Flash MiMo-V2.6-Flash none Release: 2026-09-22
Score 5.4 5.4
Rank #263 #258
Reliability 10.0 10.0
Consistency 8.8 9.2
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 36.4% 31.8%
Flaky tests 3 2
Total Runs 66 66
Cost per result 2.357 0.615
Total Cost $0.029 $0.037
Input Price $0.200 / 1M $0.140 / 1M
Output Price $1.200 / 1M $0.280 / 1M
Total Input Tokens 101,332 115,134
Output Tokens 6,709 74,105
Reasoning Tokens 0 0
Response Time (avg) 1.50s 3.64s
Response Time (max) 10.57s 55.14s
Response Time (total) 33.05s 80.06s
Parameters ~400B total (~17B active) 309B total (15B active)
Availability Closed Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#263 GPT-5.6 Luna

none
Cost
$0.016
Time
15.8s
Tokens
2,685 tok

#258 MiMo-V2.6-Flash

none
Cost
$0.008
Time
173.0s
Tokens
28,061 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
GPT-5.6 Luna 3.8 7.2 22.2% 1 980ms 7,302 459 0
MiMo-V2.6-Flash 3.7 7.6 11.1% 1 744ms 7,440 558 0

Quick Compare

Switch Comparison Pair