AI BENCHY
Advertise here
#29

GLM 5.3 Flash

Z.ai Release: 2026-08-26 Tested on: 2026-08-26 17:18 z-ai/glm-5.3-flash::max
320B total (18B active)MoEOpen source
(max) (high) (low)

Summary

GLM 5.3 Flash scores 8.7 on AI BENCHY and ranks #29. It has 10.0 reliability, a 81.8% pass rate, $0.059 total cost, and 52.10s average response time.

What makes GLM 5.3 Flash unique: Its total benchmark cost is unusually low for its score range.

Model facts

Researched on 2026-08-26

Reported
Parameters
320B total (18B active)
Architecture
MoE
Availability
Open source
License
MIT

Z.ai reports a hybrid sparse and linear attention architecture for this native multimodal checkpoint.

Score

8.7

Consistency

8.9

Total Output Tokens

204,340

Total Input Tokens

96,873

Input Price

$0.075 / 1M

Output Price

$0.250 / 1M

Tests Correct

Wrong Tests: 6

Attempt pass rate: 81.8%

Flaky tests

3

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

52.10s

Response Time (max): 333.47s

Response Time (total): 1146.24s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#29 GLM 5.3 Flash

max
Cost
$0.005
Time
296.5s
Tokens
16,250 tok

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Anti-AI Tricks 10.0 10.0
Coding 10.0 10.0
Combined 7.3 5.8
Data parsing and extraction 10.0 10.0
Domain specific 3.6 7.2
General Intelligence 10.0 10.0
Instructions following 9.8 10.0
Puzzle Solving 8.7 7.7
Tool Calling 10.0 10.0
Trivia 3.0 10.0

Compared models