Navigate
Advertise here

Gemini 3.8 Flash (low) vs GLM 5 (medium)

The average score is effectively tied at 7.2 vs 7.2. Gemini 3.8 Flash (low) has the lower benchmark cost at $0.323 vs $0.360. Gemini 3.8 Flash (low) is faster at 7.54s vs 39.05s, with pass rates of 81.2% vs 76.8%.

Last updated at: 2026-10-01

Compared models

Rank
#142
Total Output Tokens
42,741
Response Time (avg)
7.54s
Total Cost
$0.323
Rank
#146
Total Output Tokens
138,257
Response Time (avg)
39.05s
Total Cost
$0.360
Recommended model Gemini 3.8 Flash (low)

It has the best score here (7.2), while responding about 5.2x faster than GLM 5 (medium).

Detailed comparison

Metric Gemini 3.8 Flash Gemini 3.8 Flash low Release: 2026-09-02 GLM 5 GLM 5 medium Release: 2026-02-12
Score 7.2 7.2
Rank #142 #146
Reliability 9.9 10.0
Consistency 8.2 7.9
Attempts 69/69 66/69
Tests Correct
Attempt pass rate 81.2% 76.8%
Flaky tests 5 5
Total Runs 69 66
Cost per result 2.014 2.552
Total Cost $0.323 $0.360
Input Price $0.750 / 1M $0.600 / 1M
Output Price $3.750 / 1M $1.920 / 1M
Total Input Tokens 215,836 212,440
Output Tokens 9,452 22,922
Reasoning Tokens 33,289 115,335
Response Time (avg) 7.54s 39.05s
Response Time (max) 67.18s 110.70s
Response Time (total) 173.36s 546.68s
Parameters ~500B total (~20B active) 744B total (40B active)
Availability Closed Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#142 Gemini 3.8 Flash

low
Cost
$0.030
Time
45.8s
Tokens
8,023 tok

#146 GLM 5

medium
Cost
$0.005
Time
20.7s
Tokens
2,068 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.8 Flash 5.4 7.2 44.4% 1 4.31s 8,118 449 5,706
GLM 5 10.0 10.0 100.0% 0 74.30s 7,254 2,997 52,930

Quick Compare

Switch Comparison Pair