Navigate
Advertise here

Laguna XS 2.1 (medium) vs GLM 5.3 (low)

The average score is effectively tied at 6.4 vs 6.5. Laguna XS 2.1 (medium) has the lower benchmark cost at $0.083 vs $0.473. GLM 5.3 (low) is faster at 15.28s vs 48.97s, with pass rates of 46.4% vs 63.8%.

Last updated at: 2026-10-01

Compared models

Rank
#200
Total Output Tokens
538,292
Response Time (avg)
48.97s
Total Cost
$0.083
Rank
#195
Total Output Tokens
28,602
Response Time (avg)
15.28s
Total Cost
$0.473
Recommended model Laguna XS 2.1 (medium)

It has the best score here (6.4), while costing about 5.7x less than GLM 5.3 (low).

Detailed comparison

Metric Laguna XS 2.1 Laguna XS 2.1 medium Release: 2026-07-02 Free Available GLM 5.3 GLM 5.3 low Release: 2026-08-20
Score 6.4 6.5
Rank #200 #195
Reliability 10.0 10.0
Consistency 9.0 8.3
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 46.4% 63.8%
Flaky tests 3 5
Total Runs 69 69
Cost per result 0.922 3.936
Total Cost $0.083 $0.473
Input Price $0.060 / 1M $1.400 / 1M
Output Price $0.120 / 1M $4.400 / 1M
Total Input Tokens 352,516 247,436
Output Tokens 33,786 11,005
Reasoning Tokens 504,506 17,597
Response Time (avg) 48.97s 15.28s
Response Time (max) 422.72s 83.60s
Response Time (total) 1126.35s 351.49s
Parameters 33B total (3B active) 744B total (40B active)
Availability Weights available Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#200 Laguna XS 2.1

medium
Cost
$0.001
Time
30.6s
Tokens
4,678 tok

#195 GLM 5.3

low
Cost
$0.007
Time
28.4s
Tokens
1,599 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Laguna XS 2.1 5.5 10.0 33.3% 0 70.35s 7,995 23,767 83,258
GLM 5.3 6.2 6.9 55.6% 1 21.02s 7,317 368 6,764

Quick Compare

Switch Comparison Pair