Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

MoonshotAI: Kimi K2.6 vs Z.ai: GLM 5.1

Last updated at: 2026-05-29

Metric Kimi K2.6 Kimi K2.6 none Release: 2026-04-20 Free Available GLM 5.1 GLM 5.1 none Release: 2026-04-07
Score 5.6 5.8
Rank #119 #108
Reliability 10.0 10.0
Consistency 9.2 8.4
Tests Correct
Attempt pass rate 38.3% 43.3%
Flaky tests 2 4
Total Runs 60 60
Cost per result 1.241 0.806
Total Cost $0.087 $0.057
Input Price $0.730 / 1M $0.980 / 1M
Output Price $3.490 / 1M $3.080 / 1M
Output Tokens 16,405 3,748
Reasoning Tokens 0 0
Response Time (avg) 13.82s 4.20s
Response Time (max) 238.89s 32.57s
Response Time (total) 276.39s 83.95s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 4.6 10.0 25.0% 0 1.39s 471 0
GLM 5.1 4.0 6.3 25.0% 2 2.11s 305 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 7.1 9.8 50.0% 0 122.77s 14,749 0
GLM 5.1 4.3 9.5 0.0% 0 6.33s 519 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 3.0 10.0 0.0% 0 3.38s 290 0
GLM 5.1 2.8 2.1 33.3% 1 32.57s 2,129 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 10.0 10.0 100.0% 0 1.32s 201 0
GLM 5.1 10.0 10.0 100.0% 0 1.08s 204 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 5.3 7.2 44.4% 1 1.48s 42 0
GLM 5.1 2.9 7.2 11.1% 1 1.99s 24 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 5.4 3.5 33.3% 1 1.55s 138 0
GLM 5.1 5.0 10.0 0.0% 0 790ms 39 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 6.5 10.0 50.0% 0 1.64s 72 0
GLM 5.1 9.8 10.0 100.0% 0 1.98s 66 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 3.1 9.9 0.0% 0 1.40s 185 0
GLM 5.1 7.7 10.0 66.7% 0 1.45s 151 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 10.0 10.0 100.0% 0 4.46s 240 0
GLM 5.1 10.0 10.0 100.0% 0 10.68s 300 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 3.0 10.0 0.0% 0 1.36s 17 0
GLM 5.1 3.0 10.0 0.0% 0 2.34s 11 0

Quick Compare

Switch Comparison Pair