Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

GPT-5.2 (medium) vs Qwen3.8 Max (0902) (low)

GPT-5.2 (medium) leads on average score with 8.7 vs 8.6. Qwen3.8 Max (0902) (low) has the lower benchmark cost at $1.131 vs $1.540. Qwen3.8 Max (0902) (low) is faster at 29.07s vs 31.29s, with pass rates of 73.9% vs 85.5%.

Last updated at: 2026-10-01

Compared models

Rank
#54
Total Output Tokens
80,538
Response Time (avg)
31.29s
Total Cost
$1.540
Rank
#63
Total Output Tokens
88,923
Response Time (avg)
29.07s
Total Cost
$1.131
Recommended model Qwen3.8 Max (0902) (low)

It offers the best overall trade-off: a competitive score (8.6), lower cost than GPT-5.2 (medium), and balanced response time.

Detailed comparison

Metric GPT-5.2 GPT-5.2 medium Release: 2025-12-11 Qwen3.8 Max (0902) Qwen3.8 Max (0902) low Release: 2026-09-07
Score 8.7 8.6
Rank #54 #63
Reliability 10.0 10.0
Consistency 8.5 9.0
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 73.9% 85.5%
Flaky tests 4 3
Total Runs 69 69
Cost per result 10.261 6.283
Total Cost $1.540 $1.131
Input Price $1.750 / 1M $2.000 / 1M
Output Price $14.000 / 1M $6.000 / 1M
Total Input Tokens 235,207 298,682
Output Tokens 11,065 8,374
Reasoning Tokens 69,473 80,549
Response Time (avg) 31.29s 29.07s
Response Time (max) 116.86s 207.79s
Response Time (total) 531.98s 668.66s
Parameters ~1.2T total (~80B active) 2.4T total (~100B active)
Availability Closed Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#54 GPT-5.2

medium
Cost
$0.047
Time
49.2s
Tokens
3,396 tok

#63 Qwen3.8 Max (0902)

low
Cost
$0.014
Time
41.2s
Tokens
2,421 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
GPT-5.2 10.0 10.0 100.0% 0 22.73s 7,302 511 11,912
Qwen3.8 Max (0902) 10.0 10.0 100.0% 0 37.42s 8,127 485 17,404

Quick Compare

Switch Comparison Pair