Navigate
AI BENCHY
Advertise here

Poolside: Laguna S 2.1 vs Z.ai: GLM 4.7 Flash

Laguna S 2.1 (low) leads on average score with 5.0 vs 4.9. GLM 4.7 Flash has the lower benchmark cost at $0.016 vs $0.091. GLM 4.7 Flash is faster at 9.15s vs 85.35s, with pass rates of 25.8% vs 34.9%.

Recommended modelGLM 4.7 FlashIts score stays close to the best score here (4.9 vs 5.0), while costing about 5.9x less than Laguna S 2.1 (low).

Last updated at: 2026-07-22

Metric Laguna S 2.1 Laguna S 2.1 low Release: 2026-07-21 Free Available GLM 4.7 Flash GLM 4.7 Flash none Release: 2026-01-19
Score 5.0 4.9
Rank #181 #185
Reliability 9.9 10.0
Consistency 7.8 8.9
Tests Correct
Attempt pass rate 25.8% 34.9%
Flaky tests 6 3
Total Runs 66 66
Cost per result 3.022 0.256
Total Cost $0.091 $0.016
Input Price $0.100 / 1M $0.061 / 1M
Output Price $0.200 / 1M $0.400 / 1M
Total Input Tokens 118,746 101,504
Output Tokens 67,823 22,992
Reasoning Tokens 383,885 0
Response Time (avg) 85.35s 9.15s
Response Time (max) 814.73s 97.15s
Response Time (total) 1877.64s 137.18s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#181 Laguna S 2.1

low
Cost
$0.001
Time
6.6s
Tokens
1,314 tok

#185 GLM 4.7 Flash

none
Invalid SVG
Cost
$0.000
Time
300.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Laguna S 2.1 3.8 7.1 22.2% 1 167.51s 7,917 58,111 119,138
GLM 4.7 Flash 4.3 10.0 0.0% 0 2.54s 7,256 650 0

Quick Compare

Switch Comparison Pair