Muse Glimmer 30B (batch)
Muse Glimmer 30B · Batch tier (Batch)
Asynchronous batch tier: discounted offline batch requests with high throughput and relaxed latency — ideal for large-scale non-real-time workloads.
Best for: Large-scale offline inference, cost-sensitive batch jobs, backfills
This page covers the Muse Glimmer 30B Batch tier, sharing the underlying model capability with its real-time standard tier; the main difference is in billing mode and invocation method. See the Muse Glimmer 30B standard real-time tier.
Provider:Meta
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Key specs
- Context:131.1K
- Max output:118K
- Tokenizer:Other
- Released:2026-08-09
Token pricing
- Input price:$0.350 / 1M tokens
- Output price:$1.50 / 1M tokens
- Cache read:$0.040 / 1M tokens
- Blended price:$0.637 / 1M tokens
Modalities
Text, Image
Use Cases
- Deep Reasoning & Analysis:Math proofs, logical reasoning, multi-step planning — ideal for research and financial analysis.(≈ $0.012/call)
- Vision Understanding & Generation:Recognize data trends in charts, OCR scanned docs, analyze photos and produce structured reports.(≈ $0.0090/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0030/call)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input price: cost-effective for tasks that consume large context (RAG, long-doc analysis).
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
- Multimodal (text+image): fit for image captioning, OCR, chart QA, vision-grounded reasoning.
FAQ
- What is the token pricing of Muse Glimmer 30B (batch)? Input price $0.350 / 1M tokens, output price $1.50 / 1M tokens, blended about $0.637 / 1M tokens. Refer to Meta’s official page for the exact rate.
- What context window does Muse Glimmer 30B (batch) support? Context window is 131.1K, max output about 118K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Muse Glimmer 30B (batch) best for? Based on its capability and pricing, Muse Glimmer 30B (batch) fits best: Deep Reasoning & Analysis, Vision Understanding & Generation, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does Muse Glimmer 30B (batch) support? Supports Text, Image modality.
- Is Muse Glimmer 30B (batch) a free model? Muse Glimmer 30B (batch) is billed per token, not a free model.
- Which provider offers Muse Glimmer 30B (batch)? Muse Glimmer 30B (batch) is offered by Meta.
More models from Meta
Full info:Muse Glimmer 30B (batch)