Navigate
AI BENCHY
Advertise here

Inkling (medium) vs GLM 5.2 (high)

The average score is effectively tied at 8.0 vs 8.0. Inkling (medium) has the lower benchmark cost at $0.377 vs $1.463. Inkling (medium) is faster at 15.66s vs 69.86s, with pass rates of 75.8% vs 69.7%.

Last updated at: 2026-08-29

Rank
#57
Total Output Tokens
63,842
Response Time (avg)
15.66s
Total Cost
$0.377
Rank
#56
Total Output Tokens
364,655
Response Time (avg)
69.86s
Total Cost
$1.463
Recommended model Inkling (medium)

It has the best score here (8.0), while costing about 3.9x less than GLM 5.2 (high).

Detailed comparison

Metric Inkling Inkling medium Release: 2026-07-18 GLM 5.2 GLM 5.2 high Release: 2026-06-17
Score 8.0 8.0
Rank #57 #56
Reliability 10.0 10.0
Consistency 8.5 8.6
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 75.8% 69.7%
Flaky tests 4 3
Total Runs 66 66
Cost per result 2.551 6.849
Total Cost $0.377 $1.463
Input Price $0.950 / 1M $1.190 / 1M
Output Price $4.050 / 1M $3.740 / 1M
Total Input Tokens 124,062 83,822
Output Tokens 12,190 72,040
Reasoning Tokens 51,652 292,615
Response Time (avg) 15.66s 69.86s
Response Time (max) 85.12s 599.43s
Response Time (total) 344.52s 1536.98s
Parameters 975B total (41B active) 744B total (40B active)
Availability Open source Open source

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#57 Thinking Machines: Inkling

medium
Cost
$0.004
Time
11.2s
Tokens
852 tok

#56 GLM 5.2

high
Invalid SVG
Cost
$0.000
Time
300.0s
Tokens
0 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Inkling 8.4 7.8 77.8% 1 25.20s 7,374 525 14,457
GLM 5.2 6.4 8.6 33.3% 0 73.03s 5,124 2,302 22,546

Quick Compare

Switch Comparison Pair