Ling-3.0-flash
Provider:inclusionAI (Ant)
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Key specs
- Context:262.1K
- Max output:32.8K
- Tokenizer:Other
- Released:2026-07-23
Token pricing
- Input price:$0.021 / 1M tokens
- Output price:$0.063 / 1M tokens
- Cache read:$0.0042 / 1M tokens
- Blended price:$0.032 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.0005/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0003/turn)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Low blended price: budget-friendly for fleet-scale deployment.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
FAQ
- What is the token pricing of Ling-3.0-flash? Input price $0.021 / 1M tokens, output price $0.063 / 1M tokens, blended about $0.032 / 1M tokens. Refer to inclusionAI (Ant)’s official page for the exact rate.
- What context window does Ling-3.0-flash support? Context window is 262.1K, max output about 32.8K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Ling-3.0-flash best for? Based on its capability and pricing, Ling-3.0-flash fits best: Deep Reasoning & Analysis, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does Ling-3.0-flash support? Supports Text modality.
- Is Ling-3.0-flash a free model? Ling-3.0-flash is billed per token, not a free model.
- Which provider offers Ling-3.0-flash? Ling-3.0-flash is offered by inclusionAI (Ant).
More models from inclusionAI (Ant)
Full info:Ling-3.0-flash