GLM-5.2
Pricing as of July 2026$1.40/1M input · $4.40/1M output · 1.0M context · Zhipu
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to GLM-5.2
Pricing Details
ZhipuMIT-licensed 753B-param MoE, coding-first, 1M-token context. Launched June 13 with only the GLM Coding Plan subscription; the per-token API ($1.40/$4.40, holding GLM 5.1's rate) and open weights landed ~June 16. Cached input ~$0.26/1M (reported, unconfirmed on Z.ai docs). Z.ai self-reported benchmarks: SWE-Bench Pro 62.1, Terminal-Bench 2.1 81.0, DeepSWE 46.2. No independent third-party eval yet. 3x peak surcharge 14:00-18:00 Beijing time.
Input / 1M tokens
$1.4
Output / 1M tokens
$4.4
Context Window
1.0M
Max Output
131K
Price History
Launched at current price on 2026-06-13. No price changes recorded.
Frequently Asked Questions
Common questions about GLM-5.2 pricing and usage
Read More About GLM-5.2
Claude Opus 4.8 just slipped off the LLM value frontier. A $3 model scores higher on the intelligence index and costs a quarter less to run, and it is not the only flagship the math has passed.
July 23, 2026 · 9 min read
Mira Murati's lab open-sourced a frontier-class model. The strange part is how hard it is to just buy tokens for it.
July 20, 2026 · 7 min read
American teams now route a third of their tokens to Chinese models to save money. The plot twist in July's pricing is that China's best model costs as much as the US flagships it beats.
July 18, 2026 · 8 min read