GLM Latest
Provider:Z.ai (Zhipu)
This model always redirects to the latest GLM model from Z.ai.
Key specs
- Context:1.3M
- Max output:131.1K
- Tokenizer:Router
- Released:2026-08-19
Token pricing
- Input price:$1.19 / 1M tokens
- Output price:$4.18 / 1M tokens
- Cache read:$0.247 / 1M tokens
- Blended price:$1.94 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.033/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0084/call)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Long context (>=1M tokens): well-suited for long-doc QA, code-base analysis, multi-turn agents.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
FAQ
- What is the token pricing of GLM Latest? Input price $1.19 / 1M tokens, output price $4.18 / 1M tokens, blended about $1.94 / 1M tokens. Refer to Z.ai (Zhipu)’s official page for the exact rate.
- What context window does GLM Latest support? Context window is 1.3M, max output about 131.1K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is GLM Latest best for? Based on its capability and pricing, GLM Latest fits best: Deep Reasoning & Analysis, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does GLM Latest support? Supports Text modality.
- Is GLM Latest a free model? GLM Latest is billed per token, not a free model.
- Which provider offers GLM Latest? GLM Latest is offered by Z.ai (Zhipu).
More models from Z.ai (Zhipu)
Full info:GLM Latest