Mercury 2
Provider:Inception Labs
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Key specs
- Context:128K
- Max output:50K
- Tokenizer:Other
- Released:2026-03-04
Token pricing
- Input price:$0.250 / 1M tokens
- Output price:$0.750 / 1M tokens
- Cache read:$0.025 / 1M tokens
- Blended price:$0.375 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Math proofs, logical reasoning, multi-step planning — ideal for research and financial analysis.(≈ $0.0060/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0037/turn)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
FAQ
- What is the token pricing of Mercury 2? Input price $0.250 / 1M tokens, output price $0.750 / 1M tokens, blended about $0.375 / 1M tokens. Refer to Inception Labs’s official page for the exact rate.
- What context window does Mercury 2 support? Context window is 128K, max output about 50K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Mercury 2 best for? Based on its capability and pricing, Mercury 2 fits best: Deep Reasoning & Analysis, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does Mercury 2 support? Supports Text modality.
- Is Mercury 2 a free model? Mercury 2 is billed per token, not a free model.
- Which provider offers Mercury 2? Mercury 2 is offered by Inception Labs.
Full info:Mercury 2