- Rank
- #201
- Total Output Tokens
- 96,943
- Response Time (avg)
- 4.40s
- Total Cost
- $0.121
Mercury 2 (medium) vs GLM 5.3 (low)
Mercury 2 (medium) leads on average score with 6.6 vs 6.5. Mercury 2 (medium) has the lower benchmark cost at $0.121 vs $0.215. Mercury 2 (medium) is faster at 4.40s vs 15.18s, with pass rates of 53.6% vs 63.8%.
Compared models
- Rank
- #208
- Total Output Tokens
- 28,106
- Response Time (avg)
- 15.18s
- Total Cost
- $0.215
Recommended model
Mercury 2 (medium)
It has the best score here (6.6), while costing about 1.8x less than GLM 5.3 (low).
Detailed comparison
| Metric | Mercury 2 Mercury 2 medium | GLM 5.3 GLM 5.3 low |
|---|---|---|
| Score | 6.6 | 6.5 |
| Rank | #201 | #208 |
| Reliability | 10.0 | 10.0 |
| Consistency | 8.1 | 8.3 |
| Attempts | 69/69 | 69/69 |
| Tests Correct | ||
| Attempt pass rate | 53.6% | 63.8% |
| Flaky tests | 5 | 5 |
| Total Runs | 69 | 69 |
| Cost per result | 1.201 | 3.923 |
| Total Cost | $0.121 | $0.215 |
| Input Price | $0.250 / 1M | $0.070 / 1M |
| Output Price | $0.750 / 1M | $7.000 / 1M |
| Cache Read Price | $0.025 / 1M | $0.065 / 1M |
| Cache Write Price | N/A | N/A |
| Total Input Tokens | 189,300 | 247,871 |
| Output Tokens | 10,854 | 10,791 |
| Reasoning Tokens | 86,089 | 17,315 |
| Response Time (avg) | 4.40s | 15.18s |
| Response Time (max) | 34.92s | 83.60s |
| Response Time (total) | 96.81s | 349.06s |
| Parameters | ~100B | 744B total (40B active) |
| Availability | Closed | Closed |
Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#201 Mercury 2
medium- Cost
- $0.002
- Time
- 2.1s
- Tokens
- 1,702 tok
#208 GLM 5.3
low- Cost
- $0.007
- Time
- 28.4s
- Tokens
- 1,599 tok
Category Breakdown
Quick Compare
Switch Comparison Pair
Mercury 2mediumvsStep 3.7 FlashhighMercury 2.5highvsGLM 5.3lowMercury 2.5 PreviewhighvsGLM 5.3lowKAT-Coder-Pro V2.5highvsGLM 5.3lowGrok 4.3mediumvsGLM 5.3lowMercury 2mediumvsLongCat 2.0lowMercury 2mediumvsKAT-Coder-Pro V2.5lowQwen3.6 35B A3BmediumvsGLM 5.3lowLaguna XS 2.1mediumFree AvailablevsGLM 5.3lowQwen3.5-122B-A10BnonevsGLM 5.3lowClaude Sonnet 5.5mediumvsGLM 5.3lowGemini 3 Flash PreviewnonevsGLM 5.3low