GLM 4.6V
Provider:Z.ai (Zhipu)
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
Key specs
- Context:131.1K
- Max output:32.8K
- Tokenizer:Other
- Released:2025-12-08
Token pricing
- Input price:$0.300 / 1M tokens
- Output price:$0.900 / 1M tokens
- Cache read:$0.055 / 1M tokens
- Blended price:$0.450 / 1M tokens
Modalities
Image, Text, Video
Use Cases
- Deep Reasoning & Analysis:Math proofs, logical reasoning, multi-step planning — ideal for research and financial analysis.(≈ $0.0072/call)
- Vision Understanding & Generation:Recognize data trends in charts, OCR scanned docs, analyze photos and produce structured reports.(≈ $0.0054/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0045/turn)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
- Multimodal (text+image): fit for image captioning, OCR, chart QA, vision-grounded reasoning.
- Video input support: fit for video understanding, frame QA, content moderation.
FAQ
- What is the token pricing of GLM 4.6V? Input price $0.300 / 1M tokens, output price $0.900 / 1M tokens, blended about $0.450 / 1M tokens. Refer to Z.ai (Zhipu)’s official page for the exact rate.
- What context window does GLM 4.6V support? Context window is 131.1K, max output about 32.8K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is GLM 4.6V best for? Based on its capability and pricing, GLM 4.6V fits best: Deep Reasoning & Analysis, Vision Understanding & Generation, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does GLM 4.6V support? Supports Image, Text, Video modality.
- Is GLM 4.6V a free model? GLM 4.6V is billed per token, not a free model.
- Which provider offers GLM 4.6V? GLM 4.6V is offered by Z.ai (Zhipu).
More models from Z.ai (Zhipu)
Full info:GLM 4.6V