Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Google: Gemini 3.1 Flash Lite Preview vs Z.ai: GLM 5

Summary

Gemini 3.1 Flash Lite Preview vs GLM 5 benchmark comparison: GLM 5 leads on average score with 8.6 vs 6.5. Gemini 3.1 Flash Lite Preview has the lower benchmark cost at $0.026 vs $0.228. Gemini 3.1 Flash Lite Preview is faster at 2.77s vs 33.54s, with pass rates of 61.9% vs 82.5%.

Recommended model: Gemini 3.1 Flash Lite Preview - It offers the best overall trade-off: a competitive score (6.5), lower cost than GLM 5, and balanced response time.

Last updated at: 2026-06-12

Metric Gemini 3.1 Flash Lite Preview Gemini 3.1 Flash Lite Preview low Release: 2026-03-03 GLM 5 GLM 5 medium Release: 2026-02-12
Score 6.5 8.6
Rank #81 #18
Reliability 10.0 10.0
Consistency 10.0 8.5
Tests Correct
Attempt pass rate 61.9% 82.5%
Flaky tests 0 4
Total Runs 63 63
Cost per result 0.196 1.668
Total Cost $0.026 $0.228
Input Price $0.250 / 1M $0.600 / 1M
Output Price $1.500 / 1M $1.920 / 1M
Total Input Tokens 32,715 35,224
Output Tokens 2,286 21,570
Reasoning Tokens 9,166 102,996
Response Time (avg) 2.77s 33.54s
Response Time (max) 11.91s 99.85s
Response Time (total) 58.12s 435.99s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#81 Gemini 3.1 Flash Lite Preview

low
Cost
$0.002
Time
3.7s
Tokens
1,203 tok

#18 GLM 5

medium
Cost
$0.005
Time
20.7s
Tokens
2,068 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 8.3 10.0 75.0% 0 2.12s 506 462 1,638
GLM 5 10.0 10.0 100.0% 0 23.66s 555 480 7,056
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 5.5 10.0 33.3% 0 1.39s 8,138 660 1,060
GLM 5 10.0 10.0 100.0% 0 74.30s 7,254 2,997 52,930
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 11.91s 8,381 225 762
GLM 5 10.0 10.0 100.0% 0 28.96s 12,804 662 3,242
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 3.00s 7,455 291 696
GLM 5 7.1 5.6 83.3% 1 8.90s 5,508 567 3,734
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 5.3 10.0 33.3% 0 2.36s 641 18 1,212
GLM 5 3.5 4.4 33.3% 2 0ms 260 13,176 14,137
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 4.0 10.0 0.0% 0 1.54s 490 69 384
GLM 5 6.1 3.1 66.7% 1 14.69s 477 2,020 2,248
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.49s 621 72 753
GLM 5 10.0 10.0 100.0% 0 7.25s 636 1,001 2,129
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.69s 566 243 1,248
GLM 5 10.0 10.0 100.0% 0 11.33s 609 33 4,076
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 9.54s 5,757 237 993
GLM 5 10.0 10.0 100.0% 0 15.93s 6,935 233 994
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 1.35s 160 9 420
GLM 5 3.0 10.0 0.0% 0 67.37s 186 401 12,450

Quick Compare

Switch Comparison Pair