Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

OpenAI: GPT-5.5 vs Hunter Alpha

GPT-5.5 (low) leads on average score with 9.3 vs 4.7. Hunter Alpha (medium) has the lower benchmark cost at $0.000 vs $1.253. GPT-5.5 (low) is faster at 10.13s vs 10.33s, with pass rates of 86.4% vs 53.0%.

Recommended modelGPT-5.5 (low)It has the strongest score in this comparison (9.3) and the best overall balance of cost and response time across all 2 models.

Last updated at: 2026-07-25

Metric GPT-5.5 GPT-5.5 low Release: 2026-04-24 Hunter Alpha Hunter Alpha medium Release: 2026-03-11
Score 9.3 4.7
Rank #9 #199
Reliability 10.0 N/A
Consistency 10.0 6.0
Tests Correct
Attempt pass rate 86.4% 53.0%
Flaky tests 0 6
Total Runs 66 52
Cost per result 6.594 0.000
Total Cost $1.253 $0.000
Input Price $5.000 / 1M $0.000 / 1M
Output Price $30.000 / 1M $0.000 / 1M
Total Input Tokens 80,058 28,927
Output Tokens 5,378 4,682
Reasoning Tokens 23,040 17,969
Response Time (avg) 10.13s 10.33s
Response Time (max) 56.19s 30.53s
Response Time (total) 222.82s 175.58s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#9 GPT-5.5

low
Cost
$0.068
Time
37.0s
Tokens
2,339 tok

#199 Hunter Alpha

medium
Hunter Alpha was a stealth model revealed on March 18th as an early testing version of MiMo-V2-Pro. Find it here: https://openrouter.ai/xiaomi/mimo-v2-pro
Cost
$0.000
Time
0.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
GPT-5.5 10.0 10.0 100.0% 0 15.04s 7,302 423 6,402
Hunter Alpha 9.8 3.3 0.0% 0 0ms 0 0 0

Quick Compare

Switch Comparison Pair