AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com
#258

Qwen3.8 Flash Next

Qwen Release: 2026-10-04 Tested on: 2026-10-05 01:34 qwen/qwen3.8-flash-next::none
180B total (6B active)MoEWeights available

Summary

Qwen3.8 Flash Next scores 5.8 on AI BENCHY and ranks #258. It has 10.0 reliability, a 43.5% pass rate, ~$0.011 total cost, and 5.71s average response time.

Model facts

Researched on 2026-10-04

Reported
Parameters
180B total (6B active)
Architecture
MoE
Availability
Weights available
License
Qwen Community License 1.0

Qwen reports a 125B MoE language core with 6B activated, plus 51B n-gram embeddings and 4B MTP. The 180B total sums those components; 6B is the activated language-core count, not all lookup/draft work. Public weights use the custom Qwen Community License 1.0. This benchmark uses the ISTA-DASLab GSQ-RCO IQ3_S GGUF through Strata 0.1.39 on a local RTX 3090, with vision disabled and a 65,536-token runtime context. The compression is a quantized version of the exact public checkpoint, not the managed Qwen3.8-Flash API model.

Score

5.8

Consistency

9.0

Total Cost (Current Price)

~$0.011

Total Output Tokens

9,077

Total Input Tokens

266,004

Estimated benchmark electricity cost

~$0.011

Tests Correct

Wrong Tests: 15

Attempt pass rate: 43.5%

Flaky tests

3

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

5.71s

Response Time (max): 65.25s

Response Time (total): 131.36s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#258 Qwen3.8 Flash Next

none
Cost
~$0.001
Time
27.5s
Tokens
2,983 tok

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Agentic 5.0 10.0
Anti-AI Tricks 6.5 10.0
Coding 5.5 10.0
Combined 3.0 10.0
Data parsing and extraction 10.0 10.0
Domain specific 3.0 10.0
General Intelligence 6.1 3.1
Instructions following 7.1 5.8
Puzzle Solving 6.3 7.8
Tool Calling 10.0 10.0
Trivia 3.0 10.0

Compared models