AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com
#68

Qwen3.8 27B

Qwen Release: 2026-08-14 Tested on: 2026-08-14 23:33 qwen/qwen3.8-27b::medium
27.3BOtherWeights available
(high) (medium) (low) (none)

Summary

Qwen3.8 27B scores 7.8 on AI BENCHY and ranks #68. It has 9.6 reliability, a 66.7% pass rate, N/A total cost, and 33.05s average response time.

What makes Qwen3.8 27B unique: It stands out most in Coding, where it ranks #1, while Data parsing and extraction is its weakest area at #16.

Model facts

Researched on 2026-08-14

Reported
Parameters
27.3B
Architecture
Other
Availability
Weights available
License
Apache-2.0

Exact local Q4_K_M GGUF reports 27.3B parameters and a hybrid attention/SSM architecture through Ollama model details. Tested locally on an NVIDIA GeForce RTX 3090.

Score

7.8

Consistency

9.7

Total Cost (Current Price)

N/A

Total Output Tokens

136,162

Total Input Tokens

98,266

Input Price

N/A

Output Price

N/A

Tests Correct

Wrong Tests: 8

Attempt pass rate: 66.7%

Flaky tests

1

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

33.05s

Response Time (max): 214.63s

Response Time (total): 694.04s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#68 Qwen3.8 27B

medium
Cost
N/A
Time
43.4s
Tokens
2,956 tok

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Anti-AI Tricks 10.0 10.0
Coding 10.0 10.0
Combined 8.7 6.9
Data parsing and extraction 6.5 10.0
Domain specific 5.3 10.0
General Intelligence 5.0 10.0
Instructions following 10.0 10.0
Puzzle Solving 8.3 10.0
Tool Calling 3.0 10.0
Trivia 3.0 10.0

Compared models