Navigate
AI BENCHY
Advertise here

Solar Pro 4 (xhigh) vs GLM 5 (medium)

The average score is effectively tied at 7.7 vs 7.7. Solar Pro 4 (xhigh) has the lower benchmark cost at $0.049 vs $0.307. GLM 5 (medium) is faster at 33.54s vs 115.81s, with pass rates of 65.2% vs 78.8%.

Last updated at: 2026-08-11

Rank
#63
Total Output Tokens
379,812
Response Time (avg)
115.81s
Total Cost
$0.049
Rank
#60
Total Output Tokens
124,566
Response Time (avg)
33.54s
Total Cost
$0.307
Recommended model Solar Pro 4 (xhigh)

It has the best score here (7.7), while costing about 6.3x less than GLM 5 (medium).

Detailed comparison

Metric Solar Pro 4 Solar Pro 4 xhigh Release: 2026-08-11 GLM 5 GLM 5 medium Release: 2026-02-12
Score 7.7 7.7
Rank #63 #60
Reliability 10.0 10.0
Consistency 9.4 8.1
Benchmark coverage 22/22 tests · 66/66 attempts 21/22 tests · 63/66 attempts
Tests Correct
Attempt pass rate 65.2% 78.8%
Flaky tests 2 4
Total Runs 66 63
Cost per result 0.374 1.668
Total Cost $0.049 $0.307
Input Price $0.030 / 1M $0.950 / 1M
Output Price $0.120 / 1M $2.551 / 1M
Total Input Tokens 98,107 35,224
Output Tokens 21,723 21,570
Reasoning Tokens 358,089 102,996
Response Time (avg) 115.81s 33.54s
Response Time (max) 419.47s 99.85s
Response Time (total) 2547.92s 435.99s

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#63 Solar Pro 4

xhigh
No showcase result has been generated for this model yet.
Cost
$0.000
Time
-
Tokens
0 tok

#60 GLM 5

medium
Cost
$0.005
Time
20.7s
Tokens
2,068 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Solar Pro 4 5.8 9.9 33.3% 0 292.21s 7,887 15,575 115,426
GLM 5 10.0 10.0 100.0% 0 74.30s 7,254 2,997 52,930

Quick Compare

Switch Comparison Pair