Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Google: Gemma 4 31B vs Z.ai: GLM 5

Last updated at: 2026-05-10

Metric Gemma 4 31B Gemma 4 31B medium Release: 2026-04-02 Free Available GLM 5 GLM 5 none Release: 2026-02-12
Score 8.2 6.5
Rank #14 #80
Reliability 6.7 10.0
Consistency 9.6 9.7
Tests Correct
Attempt pass rate 77.2% 49.1%
Flaky tests 1 1
Total Runs 57 57
Cost per result 0.158 0.219
Total Cost $0.023 $0.020
Input Price $0.130 / 1M $0.600 / 1M
Output Price $0.380 / 1M $1.920 / 1M
Output Tokens 14,426 1,972
Reasoning Tokens 37,964 0
Response Time (avg) 28.72s 4.18s
Response Time (max) 90.14s 11.07s
Response Time (total) 488.27s 50.12s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 10.0 10.0 100.0% 0 12.89s 962 2,046
GLM 5 4.8 10.0 25.0% 0 2.37s 275 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 4.7 1.6 66.7% 1 70.97s 3,166 5,449
GLM 5 5.6 3.5 33.3% 1 8.84s 408 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 3.0 10.0 0.0% 0 0ms 0 0
GLM 5 3.0 10.0 0.0% 0 4.98s 406 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 10.0 10.0 100.0% 0 21.11s 1,822 2,951
GLM 5 10.0 10.0 100.0% 0 5.78s 203 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 7.7 10.0 66.7% 0 38.48s 4,349 8,985
GLM 5 3.0 10.0 0.0% 0 2.24s 19 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 10.0 10.0 100.0% 0 9.57s 105 888
GLM 5 10.0 10.0 100.0% 0 3.27s 103 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 10.0 10.0 100.0% 0 12.76s 533 2,035
GLM 5 10.0 10.0 100.0% 0 1.48s 61 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 9.9 10.0 100.0% 0 27.63s 1,797 5,596
GLM 5 7.7 10.0 66.7% 0 2.05s 264 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 3.0 10.0 0.0% 0 0ms 0 0
GLM 5 10.0 10.0 100.0% 0 11.07s 220 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemma 4 31B 3.0 10.0 0.0% 0 90.14s 1,692 10,014
GLM 5 3.0 10.0 0.0% 0 3.62s 13 0

Quick Compare

Switch Comparison Pair