Navigate
Advertise here

Qwen3.8 27B (high) vs Qwen3.8 Flash Next (xhigh)

Qwen3.8 27B (high) leads on average score with 8.7 vs 7.1. Qwen3.8 Flash Next (xhigh) has the lower benchmark cost at ~$0.195 vs ~$0.298. Qwen3.8 Flash Next (xhigh) is faster at 102.66s vs 166.34s, with pass rates of 78.3% vs 72.5%.

Last updated at: 2026-10-05

Compared models

Rank
#55
Total Output Tokens
677,080
Response Time (avg)
166.34s
Total Cost
~$0.298
Rank
#154
Total Output Tokens
621,056
Response Time (avg)
102.66s
Total Cost
~$0.195
Recommended model Qwen3.8 27B (high)

It has the strongest score in this comparison (8.7) and the best overall balance of cost and response time across all 2 models.

Detailed comparison

Metric Qwen3.8 27B Qwen3.8 27B high Release: 2026-08-14 Free Available Qwen3.8 Flash Next Qwen3.8 Flash Next xhigh Release: 2026-10-04
Score 8.7 7.1
Rank #55 #154
Reliability 9.7 9.9
Consistency 9.2 7.9
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 78.3% 72.5%
Flaky tests 2 6
Total Runs 69 69
Cost per result ~1.751 ~1.390
Total Cost ~$0.298 ~$0.195
Input Price N/A N/A
Output Price N/A N/A
Cache Read Price N/A N/A
Cache Write Price N/A N/A
Total Input Tokens 348,967 301,749
Output Tokens 1,005 1,033
Reasoning Tokens 676,075 620,023
Response Time (avg) 166.34s 102.66s
Response Time (max) 640.85s 674.08s
Response Time (total) 3825.81s 2361.10s
Parameters 27.3B 180B total (6B active)
Availability Weights available Weights available

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#55 Qwen3.8 27B

high
Cost
~$0.008
Time
270.1s
Tokens
18,963 tok

#154 Qwen3.8 Flash Next

xhigh
Cost
~$0.010
Time
333.1s
Tokens
30,654 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Qwen3.8 27B 8.1 7.0 88.9% 1 248.62s 8,235 346 143,335
Qwen3.8 Flash Next 7.4 7.0 77.8% 1 192.20s 8,235 122 186,070

Quick Compare

Switch Comparison Pair

GPT-5.2mediumvsQwen3.8 27BhighFree AvailableQwen3.8 27BhighFree AvailablevsMiMo-V2.6-PromediumSeed 2.1 TurbomediumvsQwen3.8 27BhighFree AvailableGPT-6 SollowvsQwen3.8 27BhighFree AvailableQwen3.8 Flash NextxhighvsGLM 5.3 FlashXmediumQwen3.8 Flash NextxhighvsMiMo-V2.6-FlashlowClaude Opus 5nonevsQwen3.8 Flash NextxhighQwen3.8 Flash NextxhighvsMiMo-V2.5mediumDeepSeek V4.1 FlashmaxvsQwen3.8 27BhighFree AvailableMuse Spark 1.2lowvsQwen3.8 27BhighFree AvailableQwen3.8 Flash NextxhighvsStep 3.7 FlashlowDeepSeek V4.1 FlashmediumvsQwen3.8 27BhighFree Available