GLM-5.2
Last verified: July 20, 2026$1.40/1M input · $4.40/1M output · 1.0M context · Zhipu
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to GLM-5.2
Pricing Details
ZhipuMIT-licensed 753B-param MoE, coding-first, 1M-token context. Launched June 13 with only the GLM Coding Plan subscription; the per-token API ($1.40/$4.40, holding GLM 5.1's rate) and open weights landed ~June 16. Cached input ~$0.26/1M (reported, unconfirmed on Z.ai docs). Z.ai self-reported benchmarks: SWE-Bench Pro 62.1, Terminal-Bench 2.1 81.0, DeepSWE 46.2. No independent third-party eval yet. 3x peak surcharge 14:00-18:00 Beijing time.
Input / 1M tokens
$1.4
Output / 1M tokens
$4.4
Context Window
1.0M
Max Output
131K
Price History
Launched at current price on 2026-06-13. No price changes recorded.
Frequently Asked Questions
Common questions about GLM-5.2 pricing and usage
Read More About GLM-5.2
Mira Murati's lab open-sourced a frontier-class model. The strange part is how hard it is to just buy tokens for it.
July 20, 2026 · 7 min read
American teams now route a third of their tokens to Chinese models to save money. The plot twist in July's pricing is that China's best model costs as much as the US flagships it beats.
July 18, 2026 · 8 min read
A Kuaishou team just priced 94 percent of Opus 4.8's coding score at a seventh of the cost. The two things holding KAT-Coder v2.5 back are that you can't download it and it stumbles in the terminal.
July 16, 2026 · 8 min read