Qwen3 32B
Provider:Qwen
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Key specs
- Context:131.1K
- Max output:16.4K
- Tokenizer:Qwen3
- Released:2025-04-28
Token pricing
- Input price:$0.080 / 1M tokens
- Output price:$0.280 / 1M tokens
- Blended price:$0.130 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Math proofs, logical reasoning, multi-step planning — ideal for research and financial analysis.(≈ $0.0022/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0014/turn)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Low blended price: budget-friendly for fleet-scale deployment.
FAQ
- What is the token pricing of Qwen3 32B? Input price $0.080 / 1M tokens, output price $0.280 / 1M tokens, blended about $0.130 / 1M tokens. Refer to Qwen’s official page for the exact rate.
- What context window does Qwen3 32B support? Context window is 131.1K, max output about 16.4K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Qwen3 32B best for? Based on its capability and pricing, Qwen3 32B fits best: Deep Reasoning & Analysis, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does Qwen3 32B support? Supports Text modality.
- Is Qwen3 32B a free model? Qwen3 32B is billed per token, not a free model.
- Which provider offers Qwen3 32B? Qwen3 32B is offered by Qwen.
More models from Qwen
Full info:Qwen3 32B