gpt-oss-120b (batch)

gpt-oss-120b · Batch tier (Batch)

Asynchronous batch tier: discounted offline batch requests with high throughput and relaxed latency — ideal for large-scale non-real-time workloads.

Best for: Large-scale offline inference, cost-sensitive batch jobs, backfills

This page covers the gpt-oss-120b Batch tier, sharing the underlying model capability with its real-time standard tier; the main difference is in billing mode and invocation method. See the gpt-oss-120b standard real-time tier.

Provider:OpenAI

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Key specs

Token pricing

Modalities

Text

Use Cases

When to pick this model

Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.

FAQ

More models from OpenAI

Full info:gpt-oss-120b (batch)