Qwen3.8 2.4T A95B (batch)
Qwen3.8 2.4T A95B · Batch tier (Batch)
Asynchronous batch tier: discounted offline batch requests with high throughput and relaxed latency — ideal for large-scale non-real-time workloads.
Best for: Large-scale offline inference, cost-sensitive batch jobs, backfills
This page covers the Qwen3.8 2.4T A95B Batch tier, sharing the underlying model capability with its real-time standard tier; the main difference is in billing mode and invocation method. See the Qwen3.8 2.4T A95B standard real-time tier.
Provider:Qwen
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Key specs
- Context:1M
- Max output:909K
- Tokenizer:Qwen
- Released:2026-08-12
Token pricing
- Input price:$2.50 / 1M tokens
- Output price:$6.25 / 1M tokens
- Cache read:$0.500 / 1M tokens
- Blended price:$3.44 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.050/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.013/call)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Long context (>=1M tokens): well-suited for long-doc QA, code-base analysis, multi-turn agents.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
FAQ
- What is the token pricing of Qwen3.8 2.4T A95B (batch)? Input price $2.50 / 1M tokens, output price $6.25 / 1M tokens, blended about $3.44 / 1M tokens. Refer to Qwen’s official page for the exact rate.
- What context window does Qwen3.8 2.4T A95B (batch) support? Context window is 1M, max output about 909K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Qwen3.8 2.4T A95B (batch) best for? Based on its capability and pricing, Qwen3.8 2.4T A95B (batch) fits best: Deep Reasoning & Analysis, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does Qwen3.8 2.4T A95B (batch) support? Supports Text modality.
- Is Qwen3.8 2.4T A95B (batch) a free model? Qwen3.8 2.4T A95B (batch) is billed per token, not a free model.
- Which provider offers Qwen3.8 2.4T A95B (batch)? Qwen3.8 2.4T A95B (batch) is offered by Qwen.
More models from Qwen
Full info:Qwen3.8 2.4T A95B (batch)