zai-glm-5-2 · Z.ai
Z.ai's GLM 5.2 at $1.40/$4.40 — a frontier-adjacent open-weight line, verified at Mistral's official API rates.
Last checked Aug 16, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Note: Z.ai's own billing is credit-multiplier based; USD rate shown is Mistral's official hosted price. Context/max-output not published — conservative defaults shown. Confirm the current rate at the official provider source.
Input
$1.40
per 1M tokens
Output
$4.40
per 1M tokens
Cached input
$0.14
per 1M tokens
Context window
200K
max output 32.8K
GLM 5.2 cost calculator
What GLM 5.2 costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.003353 | $100.59 |
| RAG / search-augmented answers | 8,000 | 500 | $0.00836 | $250.80 |
| AI coding assistant | 12,000 | 2,000 | $0.0165 | $495.84 |
| Document summarization | 25,000 | 600 | $0.0345 | $1,034.70 |
| Agentic workflow | 40,000 | 1,500 | $0.0298 | $895.20 |
| Content generation | 800 | 1,200 | $0.006098 | $182.928 |
| Data extraction & tagging | 2,000 | 250 | $0.003396 | $101.88 |
| Translation | 5,000 | 5,500 | $0.0303 | $907.65 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| GLM 5.2 | $1.40 | $4.40 | 200K | $0.0066 |
| GPT-5.6 Terracompare | $2.00 | $12.00 | 1.1M | $0.014 |
| DeepSeek V4 Pro | $1.32 | $3.96 | 1M | $0.005896 |
| Gemini 3.5 Flash | $1.50 | $9.00 | 1M | $0.0105 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
GLM 5.2 costs $1.40 per 1M input tokens and $4.40 per 1M output tokens, with cached input at $0.14 per 1M tokens. Verified Aug 16, 2026 against https://mistral.ai/pricing/api/.
GLM 5.2 supports a 200,000-token context window with up to 32,768 output tokens per request, tokenized with GLM tokenizer.
At $1.40/M input and $4.40/M output, GLM 5.2 sits below GPT-5.6 Terra ($2.00/M in, $12.00/M out) — see the comparison table for full-workload differences.