AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com
#154

Qwen3.8 Flash Next

Qwen Release: 2026-10-04 Tested on: 2026-10-05 01:35 qwen/qwen3.8-flash-next::xhigh
180B total (6B active)MoEWeights available

Summary

Qwen3.8 Flash Next scores 7.1 on AI BENCHY and ranks #154. It has 9.9 reliability, a 72.5% pass rate, ~$0.195 total cost, and 102.66s average response time.

What makes Qwen3.8 Flash Next unique: It uses unusually many reasoning tokens, which can help explain its slower or more expensive runs.

Model facts

Researched on 2026-10-04

Reported
Parameters
180B total (6B active)
Architecture
MoE
Availability
Weights available
License
Qwen Community License 1.0

Qwen reports a 125B MoE language core with 6B activated, plus 51B n-gram embeddings and 4B MTP. The 180B total sums those components; 6B is the activated language-core count, not all lookup/draft work. Public weights use the custom Qwen Community License 1.0. This benchmark uses the ISTA-DASLab GSQ-RCO IQ3_S GGUF through Strata 0.1.39 on a local RTX 3090, with vision disabled and a 65,536-token runtime context. The compression is a quantized version of the exact public checkpoint, not the managed Qwen3.8-Flash API model.

Score

7.1

Consistency

7.9

Total Cost (Current Price)

~$0.195

Total Output Tokens

621,056

Total Input Tokens

301,749

Estimated benchmark electricity cost

~$0.195

Tests Correct

Wrong Tests: 9

Attempt pass rate: 72.5%

Flaky tests

6

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

102.66s

Response Time (max): 674.08s

Response Time (total): 2361.10s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#154 Qwen3.8 Flash Next

xhigh
Cost
~$0.010
Time
333.1s
Tokens
30,654 tok

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Agentic 4.7 3.1
Anti-AI Tricks 10.0 10.0
Coding 7.4 7.0
Combined 7.3 5.8
Data parsing and extraction 10.0 10.0
Domain specific 3.5 4.4
General Intelligence 4.7 3.1
Instructions following 9.8 10.0
Puzzle Solving 7.7 10.0
Tool Calling 10.0 10.0
Trivia 3.0 10.0

Compared models