Gemini 3.1 Flash Lite Preview
Provider:Google
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
Key specs
- Context:1M
- Max output:65.5K
- Tokenizer:Gemini
- Released:2026-03-03
Token pricing
- Input price:$0.250 / 1M tokens
- Output price:$1.50 / 1M tokens
- Cache read:$0.025 / 1M tokens
- Cache write:$0.083 / 1M tokens
- Blended price:$0.563 / 1M tokens
- Image price:$0.0003 / image
Modalities
Text, Image, Video, File, Audio
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.012/call)
- Vision Understanding & Generation:Recognize data trends in charts, OCR scanned docs, analyze photos and produce structured reports.(≈ $0.0090/call)
- Voice Interaction:Speech recognition & transcription, mixed CN/EN and dialect support — meeting notes, QA.
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input price: cost-effective for tasks that consume large context (RAG, long-doc analysis).
- Long context (>=1M tokens): well-suited for long-doc QA, code-base analysis, multi-turn agents.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
- Multimodal (text+image): fit for image captioning, OCR, chart QA, vision-grounded reasoning.
- Audio support: fit for speech-to-text / voice agent / transcript analysis.
- Video input support: fit for video understanding, frame QA, content moderation.
- File / PDF input support: fit for document parsing, contract review, RAG over private data.
FAQ
- What is the token pricing of Gemini 3.1 Flash Lite Preview? Input price $0.250 / 1M tokens, output price $1.50 / 1M tokens, blended about $0.563 / 1M tokens. Refer to Google’s official page for the exact rate.
- What context window does Gemini 3.1 Flash Lite Preview support? Context window is 1M, max output about 65.5K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Gemini 3.1 Flash Lite Preview best for? Based on its capability and pricing, Gemini 3.1 Flash Lite Preview fits best: Deep Reasoning & Analysis, Vision Understanding & Generation, Voice Interaction. See the Use Cases section for details and cost estimates.
- What input/output modalities does Gemini 3.1 Flash Lite Preview support? Supports Text, Image, Video, File, Audio modality.
- Is Gemini 3.1 Flash Lite Preview a free model? Gemini 3.1 Flash Lite Preview is billed per token, not a free model.
- Which provider offers Gemini 3.1 Flash Lite Preview? Gemini 3.1 Flash Lite Preview is offered by Google.
More models from Google
Full info:Gemini 3.1 Flash Lite Preview