Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Inkling Small (high) vs GLM 5.3 Flash (high)

Inkling Small (high) leads on average score with 7.4 vs 7.3. GLM 5.3 Flash (high) has the lower benchmark cost at $0.056 vs $0.398. Inkling Small (high) is faster at 29.10s vs 29.11s, with pass rates of 66.7% vs 66.7%.

Last updated at: 2026-10-09

Compared models

Rank
#138
Total Output Tokens
298,893
Response Time (avg)
29.10s
Total Cost
$0.398
Rank
#142
Total Output Tokens
99,340
Response Time (avg)
29.11s
Total Cost
$0.056
Recommended model GLM 5.3 Flash (high)

Its score stays close to the best score here (7.3 vs 7.4), while costing about 7.2x less than Inkling Small (high).

Detailed comparison

Metric Inkling Small Inkling Small high Release: 2026-08-01 Free Available GLM 5.3 Flash GLM 5.3 Flash high Release: 2026-08-26
Score 7.4 7.3
Rank #138 #142
Reliability 10.0 10.0
Consistency 9.3 8.6
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 66.7% 66.7%
Flaky tests 2 4
Total Runs 69 69
Cost per result 2.868 0.426
Total Cost $0.398 $0.056
Input Price $0.450 / 1M $0.150 / 1M
Output Price $1.200 / 1M $0.500 / 1M
Cache Read Price $0.100 / 1M $0.030 / 1M
Cache Write Price N/A N/A
Total Input Tokens 85,669 225,013
Output Tokens 6,092 12,197
Reasoning Tokens 292,801 87,143
Response Time (avg) 29.10s 29.11s
Response Time (max) 122.60s 163.41s
Response Time (total) 669.40s 669.54s
Parameters 276B total (12B active) 320B total (18B active)
Availability Closed Open source

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#138 Thinking Machines: Inkling Small

high
Cost
$0.006
Time
34.7s
Tokens
4,735 tok

#142 GLM 5.3 Flash

high
Cost
$0.003
Time
186.7s
Tokens
10,486 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Inkling Small 10.0 10.0 100.0% 0 93.23s 7,374 463 120,474
GLM 5.3 Flash 8.2 7.2 88.9% 1 19.62s 7,317 373 9,973

Switch Comparison Pair