stealth/ox-alpha · Z.ai (Zhipu)
Z.ai's anonymously launched GLM-family preview for coding, sustained agentic work, long-horizon software engineering, and visual context.
Last checked Aug 28, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Note: REMOVED from OpenRouter on 2026-08-28 (model id stealth/ox-alpha no longer listed). Command Code may still route the alias while the stealth preview lasts; treat availability as unconfirmed. Historical price: $0 input, $0 output during the free preview. Confirm the current rate at the official provider source.
Quick answer
At the published base rate, Ox Alpha (Z.ai GLM preview) costs $0.00 per 1M input tokens and $0.00 per 1M output tokens. It supports a 1M context window and is currently marked deprecated.
Method & trust
This rate card uses the provider's listed base input, output, and cached-input prices. Workload tables below apply those rates to explicit token counts, cache assumptions, and request volumes; verify provider-specific tiers before committing budget.
Input
Free
per 1M tokens
Output
Free
per 1M tokens
Cached input
—
not published
Context window
1M
max output 131.1K
Ox Alpha (Z.ai GLM preview) cost calculator
What Ox Alpha (Z.ai GLM preview) costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.00 | $0.00 |
| RAG / search-augmented answers | 8,000 | 500 | $0.00 | $0.00 |
| AI coding assistant | 12,000 | 2,000 | $0.00 | $0.00 |
| Document summarization | 25,000 | 600 | $0.00 | $0.00 |
| Agentic workflow | 40,000 | 1,500 | $0.00 | $0.00 |
| Content generation | 800 | 1,200 | $0.00 | $0.00 |
| Data extraction & tagging | 2,000 | 250 | $0.00 | $0.00 |
| Translation | 5,000 | 5,500 | $0.00 | $0.00 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| Ox Alpha (Z.ai GLM preview) | Free | Free | 1M | $0.00 |
| GPT-5.6 Solcompare | $5.00 | $30.00 | 1.1M | $0.035 |
| GPT-5.6 Terracompare | $2.00 | $12.00 | 1.1M | $0.014 |
| GPT-5.6 Lunacompare | $0.20 | $1.20 | 1.1M | $0.0014 |
| GPT-5.6 Cybercompare | $12.50 | $75.00 | 1.1M | $0.0875 |
| GPT-5.5 Standardcompare | $5.00 | $30.00 | 512K | $0.035 |
| GPT-5.5 Procompare | $30.00 | $180.00 | 512K | $0.21 |
| GPT-5.4 Workhorsecompare | $2.50 | $15.00 | 256K | $0.0175 |
| GPT-5.4 minicompare | $0.75 | $4.50 | 256K | $0.00525 |
*4,000 in + 800 out tokens, 50% cached input where available.
Pricing history
Related calculations
FAQ
Ox Alpha (Z.ai GLM preview) costs Free per 1M input tokens and Free per 1M output tokens. Verified Aug 28, 2026 against https://commandcode.ai/models/ox-alpha.
Ox Alpha (Z.ai GLM preview) supports a 1,048,576-token context window with up to 131,072 output tokens per request, tokenized with other.
At Free/M input and Free/M output, Ox Alpha (Z.ai GLM preview) sits below GPT-5.6 Sol ($5.00/M in, $30.00/M out) — see the comparison table for full-workload differences.