Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Claude Opus 5.5 (high) vs GPT-6.1 Sol (medium)

GPT-6.1 Sol (medium) leads on average score with 9.6 vs 9.3. GPT-6.1 Sol (medium) has the lower benchmark cost at $0.326 vs $1.266. GPT-6.1 Sol (medium) is faster at 6.99s vs 25.88s, with pass rates of 86.4% vs 93.9%.

Last updated at: 2026-09-29

Compared models

Rank
#26
Total Output Tokens
36,457
Response Time (avg)
25.88s
Total Cost
$1.266
Rank
#11
Total Output Tokens
16,852
Response Time (avg)
6.99s
Total Cost
$0.326
Recommended model GPT-6.1 Sol (medium)

It has the best score here (9.6), while costing about 3.9x less than Claude Opus 5.5 (high).

Detailed comparison

Metric Claude Opus 5.5 Claude Opus 5.5 high Release: 2026-09-23 GPT-6.1 Sol GPT-6.1 Sol medium Release: 2026-09-29
Score 9.3 9.6
Rank #26 #11
Reliability 10.0 10.0
Consistency 9.3 9.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 86.4% 93.9%
Flaky tests 2 1
Total Runs 66 66
Cost per result 7.029 1.627
Total Cost $1.266 $0.326
Input Price $4.000 / 1M $2.000 / 1M
Output Price $20.000 / 1M $10.000 / 1M
Total Input Tokens 134,020 78,368
Output Tokens 7,237 4,354
Reasoning Tokens 29,220 12,498
Response Time (avg) 25.88s 6.99s
Response Time (max) 95.95s 50.74s
Response Time (total) 569.33s 153.75s
Parameters ~5T total (~500B active) ~2T total (~150B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#26 Claude Opus 5.5

high
Cost
$0.136
Time
75.6s
Tokens
6,934 tok

#11 GPT-6.1 Sol

medium
Cost
$0.045
Time
72.7s
Tokens
4,581 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Opus 5.5 10.0 10.0 100.0% 0 13.44s 10,608 584 5,042
GPT-6.1 Sol 10.0 10.0 100.0% 0 4.74s 7,302 393 1,710

Quick Compare

Switch Comparison Pair