AI BENCHY
Advertise here
#211

Qwen3.8 27B

Qwen Release: 2026-08-14 Tested on: 2026-08-14 23:32 qwen/qwen3.8-27b::none
27.3BOtherWeights available

Summary

Qwen3.8 27B scores 5.5 on AI BENCHY and ranks #211. It has 10.0 reliability, a 40.9% pass rate, N/A total cost, and 3.12s average response time.

What makes Qwen3.8 27B unique: It stands out most in Coding, where it ranks #1, while Data parsing and extraction is its weakest area at #10. It is notably fast compared with similar models.

Model facts

Researched on 2026-08-14

Reported
Parameters
27.3B
Architecture
Other
Availability
Weights available
License
Apache-2.0

Exact local Q4_K_M GGUF reports 27.3B parameters and a hybrid attention/SSM architecture through Ollama model details. Tested locally on an NVIDIA GeForce RTX 3090.

Score

5.5

Consistency

10.0

Total Cost (Current Price)

N/A

Total Output Tokens

11,445

Total Input Tokens

122,463

Input Price

N/A

Output Price

N/A

Tests Correct

Wrong Tests: 13

Attempt pass rate: 40.9%

Flaky tests

0

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

3.12s

Response Time (max): 44.76s

Response Time (total): 68.69s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#211 Qwen3.8 27B

none
Cost
N/A
Time
23.9s
Tokens
1,922 tok

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Anti-AI Tricks 6.5 10.0
Coding 5.5 10.0
Combined 3.0 10.0
Data parsing and extraction 6.5 10.0
Domain specific 7.7 10.0
General Intelligence 4.2 9.9
Instructions following 6.3 10.0
Puzzle Solving 6.0 10.0
Tool Calling 10.0 10.0
Trivia 3.0 10.0

Compared models