Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

DeepSeek: DeepSeek V4 Flash vs Z.ai: GLM 4.7 Flash

Summary

DeepSeek V4 Flash vs GLM 4.7 Flash benchmark comparison: DeepSeek V4 Flash leads on average score with 5.0 vs 4.4. DeepSeek V4 Flash has the lower benchmark cost at $0.008 vs $0.054. DeepSeek V4 Flash is faster at 26.75s vs 35.10s, with pass rates of 30.2% vs 33.3%.

Recommended model: DeepSeek V4 Flash - It has the best score here (5.0), while costing about 7.0x less than GLM 4.7 Flash.

Last updated at: 2026-06-04

Metric DeepSeek V4 Flash DeepSeek V4 Flash none Release: 2026-04-24 GLM 4.7 Flash GLM 4.7 Flash medium Release: 2026-01-19
Score 5.0 4.4
Rank #139 #158
Reliability 10.0 6.7
Consistency 8.9 6.8
Tests Correct
Attempt pass rate 30.2% 33.3%
Flaky tests 3 8
Total Runs 63 63
Cost per result 0.203 1.337
Total Cost $0.008 $0.054
Input Price $0.099 / 1M $0.060 / 1M
Output Price $0.197 / 1M $0.400 / 1M
Total Input Tokens 50,127 37,206
Output Tokens 13,710 43,754
Reasoning Tokens 0 89,079
Response Time (avg) 26.75s 35.10s
Response Time (max) 111.96s 174.55s
Response Time (total) 561.82s 456.24s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#139 DeepSeek V4 Flash

none
Cost
$0.004
Time
157.6s
Tokens
11,297 tok

#158 GLM 4.7 Flash

medium
Invalid SVG
Cost
$0.000
Time
186.2s
Tokens
12,112 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.0 10.0 0.0% 0 20.18s 540 174 0
GLM 4.7 Flash 4.7 5.9 41.7% 2 14.95s 555 1,122 6,110
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.2 7.4 11.1% 1 17.13s 7,279 9,717 0
GLM 4.7 Flash 3.2 7.4 11.1% 1 55.33s 3,106 4,981 22,387
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.5 2.1 66.7% 1 111.96s 24,398 2,664 0
GLM 4.7 Flash 2.8 2.1 33.3% 1 65.57s 17,185 2,585 20,648
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 23.79s 7,290 195 0
GLM 4.7 Flash 6.3 10.0 50.0% 0 1.51s 7,107 584 2,755
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 5.3 10.0 33.3% 0 19.73s 666 18 0
GLM 4.7 Flash 3.5 4.4 33.3% 2 174.55s 643 33,000 25,394
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.2 9.9 0.0% 0 23.74s 471 67 0
GLM 4.7 Flash 3.6 9.7 0.0% 0 18.14s 318 18 2,138
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 6.5 10.0 50.0% 0 17.54s 627 321 0
GLM 4.7 Flash 6.2 5.8 66.7% 1 2.97s 636 388 2,181
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.1 7.3 11.1% 1 23.72s 594 207 0
GLM 4.7 Flash 2.9 7.2 11.1% 1 12.93s 521 781 5,255
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 77.93s 8,079 327 0
GLM 4.7 Flash 10.0 10.0 100.0% 0 15.95s 6,949 224 1,014
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.0 10.0 0.0% 0 3.07s 183 20 0
GLM 4.7 Flash 3.0 10.0 0.0% 0 11.13s 186 71 1,197

Quick Compare

Switch Comparison Pair