Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Claude Opus 5.5 (high) vs Gemini 3.1 Pro Preview (medium)

The average score is effectively tied at 9.3 vs 9.2. Claude Opus 5.5 (high) has the lower benchmark cost at $1.266 vs $1.352. Gemini 3.1 Pro Preview (medium) is faster at 20.62s vs 25.88s, with pass rates of 86.4% vs 90.9%.

Last updated at: 2026-09-23

Compared models

Rank
#21
Total Output Tokens
36,457
Response Time (avg)
25.88s
Total Cost
$1.266
Rank
#22
Total Output Tokens
97,238
Response Time (avg)
20.62s
Total Cost
$1.352
Recommended model Gemini 3.1 Pro Preview (medium)

It has the strongest score in this comparison (9.2) and the best overall balance of cost and response time across all 2 models.

Detailed comparison

Metric Claude Opus 5.5 Claude Opus 5.5 high Release: 2026-09-23 Gemini 3.1 Pro Preview Gemini 3.1 Pro Preview medium Release: 2026-02-19
Score 9.3 9.2
Rank #21 #22
Reliability 10.0 10.0
Consistency 9.3 10.0
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 86.4% 90.9%
Flaky tests 2 0
Total Runs 66 66
Cost per result 7.029 6.758
Total Cost $1.266 $1.352
Input Price $4.000 / 1M $2.000 / 1M
Output Price $20.000 / 1M $12.000 / 1M
Total Input Tokens 134,020 92,296
Output Tokens 7,237 5,232
Reasoning Tokens 29,220 92,006
Response Time (avg) 25.88s 20.62s
Response Time (max) 95.95s 88.68s
Response Time (total) 569.33s 329.94s
Parameters ~5T total (~500B active) ~1.2T total (~20B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#21 Claude Opus 5.5

high
Cost
$0.136
Time
75.6s
Tokens
6,934 tok

#22 Gemini 3.1 Pro Preview

medium
Cost
$0.115
Time
87.2s
Tokens
9,629 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Claude Opus 5.5 10.0 10.0 100.0% 0 13.44s 10,608 584 5,042
Gemini 3.1 Pro Preview 7.9 9.9 66.7% 0 40.17s 8,124 435 41,247

Quick Compare

Switch Comparison Pair