Nemotron 3 Nano 30B A3B
Provider:NVIDIA
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
Key specs
- Context:262.1K
- Max output:228K
- Tokenizer:Other
- Released:2025-12-14
Token pricing
- Input price:$0.050 / 1M tokens
- Output price:$0.200 / 1M tokens
- Cache read:$0.025 / 1M tokens
- Blended price:$0.087 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.0016/call)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Low blended price: budget-friendly for fleet-scale deployment.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
FAQ
- What is the token pricing of Nemotron 3 Nano 30B A3B? Input price $0.050 / 1M tokens, output price $0.200 / 1M tokens, blended about $0.087 / 1M tokens. Refer to NVIDIA’s official page for the exact rate.
- What context window does Nemotron 3 Nano 30B A3B support? Context window is 262.1K, max output about 228K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Nemotron 3 Nano 30B A3B best for? Based on its capability and pricing, Nemotron 3 Nano 30B A3B fits best: Deep Reasoning & Analysis. See the Use Cases section for details and cost estimates.
- What input/output modalities does Nemotron 3 Nano 30B A3B support? Supports Text modality.
- Is Nemotron 3 Nano 30B A3B a free model? Nemotron 3 Nano 30B A3B is billed per token, not a free model.
- Which provider offers Nemotron 3 Nano 30B A3B? Nemotron 3 Nano 30B A3B is offered by NVIDIA.
More models from NVIDIA
Full info:Nemotron 3 Nano 30B A3B