AI BENCHY Compare
OpenAI: GPT-5.4 vs Z.ai: GLM 5
Last updated at: 2026-04-16
| Metric | GPT-5.4 GPT-5.4 medium | GLM 5 GLM 5 none |
|---|---|---|
| Score | 8.2 | 6.6 |
| Rank | #16 | #52 |
| Consistency | 8.7 | 9.6 |
| Tests Correct | ||
| Attempt pass rate | 79.6% | 51.9% |
| Flaky tests | 3 | 1 |
| Total Runs | 54 | 54 |
| Cost per result | 6.399 | 0.217 |
| Total Cost | $0.832 | $0.020 |
| Input Price | $2.500 / 1M | $0.720 / 1M |
| Output Price | $15.000 / 1M | $2.300 / 1M |
| Output Tokens | 2,169 | 1,959 |
| Reasoning Tokens | 48,732 | 0 |
| Response Time (avg) | 18.63s | 4.23s |
| Response Time (max) | 100.41s | 11.07s |
| Response Time (total) | 335.26s | 46.51s |
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
Quick Compare
Switch Comparison Pair
Grok 4.1 FastmediumvsGLM 5noneNemotron 3 SupermediumFree AvailablevsGLM 5noneGemini 3 Flash PreviewnonevsGPT-5.4mediumGemini 3.1 Flash Lite PreviewlowvsGPT-5.4mediumMercury 2mediumvsGLM 5noneGemini 3.1 Flash Lite PreviewnonevsGPT-5.4mediumGrok 4.20mediumvsGLM 5noneKimi K2.5mediumvsGLM 5noneGPT-5 MinimediumvsGLM 5noneGPT-5 NanomediumvsGLM 5noneGemini 3 Flash PreviewlowvsGPT-5.4mediumGPT-5.4 MinimediumvsGLM 5none