GPT-5.6 Luna (batch)
GPT-5.6 Luna · Batch tier (Batch)
Asynchronous batch tier: discounted offline batch requests with high throughput and relaxed latency — ideal for large-scale non-real-time workloads.
Best for: Large-scale offline inference, cost-sensitive batch jobs, backfills
This page covers the GPT-5.6 Luna Batch tier, sharing the underlying model capability with its real-time standard tier; the main difference is in billing mode and invocation method. See the GPT-5.6 Luna standard real-time tier.
Provider:OpenAI
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Key specs
- Context:1.1M
- Max output:128K
- Tokenizer:GPT
- Released:2026-07-09
Token pricing
- Input price:$0.100 / 1M tokens
- Output price:$0.600 / 1M tokens
- Cache read:$0.010 / 1M tokens
- Blended price:$0.225 / 1M tokens
Modalities
File, Image, Text
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.0048/call)
- Vision Understanding & Generation:Recognize data trends in charts, OCR scanned docs, analyze photos and produce structured reports.(≈ $0.0036/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0030/turn)
- Content Moderation:Auto-detect violations, hate speech, NSFW content; custom rule sets and score thresholds.(≈ $0.0012/item)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Low blended price: budget-friendly for fleet-scale deployment.
- Long context (>=1M tokens): well-suited for long-doc QA, code-base analysis, multi-turn agents.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
- Multimodal (text+image): fit for image captioning, OCR, chart QA, vision-grounded reasoning.
- File / PDF input support: fit for document parsing, contract review, RAG over private data.
FAQ
- What is the token pricing of GPT-5.6 Luna (batch)? Input price $0.100 / 1M tokens, output price $0.600 / 1M tokens, blended about $0.225 / 1M tokens. Refer to OpenAI’s official page for the exact rate.
- What context window does GPT-5.6 Luna (batch) support? Context window is 1.1M, max output about 128K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is GPT-5.6 Luna (batch) best for? Based on its capability and pricing, GPT-5.6 Luna (batch) fits best: Deep Reasoning & Analysis, Vision Understanding & Generation, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does GPT-5.6 Luna (batch) support? Supports File, Image, Text modality.
- Is GPT-5.6 Luna (batch) a free model? GPT-5.6 Luna (batch) is billed per token, not a free model.
- Which provider offers GPT-5.6 Luna (batch)? GPT-5.6 Luna (batch) is offered by OpenAI.
More models from OpenAI
Full info:GPT-5.6 Luna (batch)