Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

DeepSeek: DeepSeek V4 Flash vs Z.ai: GLM 5 Turbo

Summary

DeepSeek V4 Flash vs GLM 5 Turbo benchmark comparison: GLM 5 Turbo leads on average score with 8.4 vs 8.3. DeepSeek V4 Flash has the lower benchmark cost at $0.029 vs $0.323. GLM 5 Turbo is faster at 23.00s vs 45.85s, with pass rates of 74.6% vs 74.6%.

Recommended model: DeepSeek V4 Flash - Its score stays close to the best score here (8.3 vs 8.4), while costing about 11.3x less than GLM 5 Turbo.

Last updated at: 2026-06-12

Metric DeepSeek V4 Flash DeepSeek V4 Flash high Release: 2026-04-24 GLM 5 Turbo GLM 5 Turbo medium Release: 2026-03-15
Score 8.3 8.4
Rank #26 #24
Reliability 10.0 10.0
Consistency 8.5 8.5
Tests Correct
Attempt pass rate 74.6% 74.6%
Flaky tests 4 4
Total Runs 63 63
Cost per result 0.299 2.011
Total Cost $0.029 $0.323
Input Price $0.098 / 1M $1.200 / 1M
Output Price $0.196 / 1M $4.000 / 1M
Total Input Tokens 39,745 35,593
Output Tokens 10,310 12,245
Reasoning Tokens 123,501 62,277
Response Time (avg) 45.85s 23.00s
Response Time (max) 218.13s 194.23s
Response Time (total) 962.79s 482.97s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#26 DeepSeek V4 Flash

high
Cost
$0.003
Time
93.1s
Tokens
7,926 tok

#24 GLM 5 Turbo

medium
Cost
$0.074
Time
206.0s
Tokens
18,549 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 8.3 10.0 75.0% 0 28.51s 540 140 7,770
GLM 5 Turbo 10.0 10.0 100.0% 0 4.82s 555 362 3,137
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 7.8 10.0 66.7% 0 50.60s 7,279 395 34,862
GLM 5 Turbo 8.2 9.3 66.7% 0 45.90s 5,941 363 25,381
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 76.57s 14,016 465 7,347
GLM 5 Turbo 10.0 10.0 100.0% 0 13.88s 12,714 390 2,037
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 28.03s 7,290 201 1,179
GLM 5 Turbo 10.0 10.0 100.0% 0 6.19s 7,107 577 3,632
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.1 4.4 44.5% 2 100.31s 666 27 59,249
GLM 5 Turbo 2.9 4.4 22.2% 2 71.07s 489 9,665 19,279
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 6.1 3.1 66.7% 1 25.15s 471 79 632
GLM 5 Turbo 6.1 3.1 66.7% 1 10.05s 477 60 2,216
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 15.36s 627 63 1,622
GLM 5 Turbo 10.0 10.0 100.0% 0 5.38s 636 255 2,183
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 8.2 7.2 88.9% 1 26.11s 594 196 1,767
GLM 5 Turbo 8.7 7.9 77.8% 1 5.23s 609 312 2,647
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 74.73s 8,079 228 542
GLM 5 Turbo 10.0 10.0 100.0% 0 9.84s 6,879 241 446
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.0 10.0 0.0% 0 54.46s 183 8,516 8,531
GLM 5 Turbo 3.0 10.0 0.0% 0 40.17s 186 20 1,319

Quick Compare

Switch Comparison Pair