gpt-oss-20b (batch)

gpt-oss-20b · Batch tier (Batch)

Asynchronous batch tier: discounted offline batch requests with high throughput and relaxed latency — ideal for large-scale non-real-time workloads.

Best for: Large-scale offline inference, cost-sensitive batch jobs, backfills

This page covers the gpt-oss-20b Batch tier, sharing the underlying model capability with its real-time standard tier; the main difference is in billing mode and invocation method. See the gpt-oss-20b standard real-time tier.

Provider:OpenAI

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Key specs

Token pricing

Modalities

Text

Use Cases

When to pick this model

Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.

FAQ

More models from OpenAI

Full info:gpt-oss-20b (batch)