Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

OpenAI: GPT-5 Mini vs Hunter Alpha

Last updated at: 2026-03-12

Metric GPT-5 Mini GPT-5 Mini medium Release: 2025-08-07 Hunter Alpha Hunter Alpha none Release: Unknown release date
Rank #34 #50
Avg Score 6.0 4.6
Consistency 8.9 8.0
Cost per result 1.457 0.000
Total Cost $0.117 $0.000
Tests Correct
Attempt pass rate 58.3% 52.1%
Flaky tests 2 4
Total Runs 48 48
Output Tokens 5,826 2,272
Reasoning Tokens 48,768 0
Response Time (avg) 25.14s 4.64s
Response Time (max) 88.15s 15.17s
Response Time (total) 402.29s 74.24s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Avg Score vs Response Time (avg)

Total Output Tokens

Avg Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 7.0 9.6 66.7% 0 16.45s 1,645 5,824
Hunter Alpha 1.3 7.4 22.2% 1 3.85s 773 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 10.0 10.0 100.0% 0 88.15s 754 11,520
Hunter Alpha 10.0 10.0 0.0% 0 15.17s 379 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 9.9 10.0 100.0% 0 12.58s 453 3,200
Hunter Alpha 9.9 10.0 100.0% 0 8.49s 249 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 10.0 7.2 22.2% 1 44.63s 293 14,016
Hunter Alpha 4.0 10.0 33.3% 0 2.33s 27 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 4.0 10.0 0.0% 0 13.50s 349 1,856
Hunter Alpha 5.0 3.1 66.7% 1 2.71s 91 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 7.5 6.6 83.3% 1 15.66s 318 4,992
Hunter Alpha 5.0 10.0 50.0% 0 2.82s 69 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 4.3 9.8 33.3% 0 14.09s 1,527 5,760
Hunter Alpha 4.0 4.4 66.7% 2 3.06s 349 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
GPT-5 Mini 10.0 10.0 100.0% 0 18.64s 487 1,600
Hunter Alpha 10.0 10.0 100.0% 0 6.02s 335 0

Quick Compare

Switch Comparison Pair