OpenAI

gpt-oss-120b

OpenAI's open-weight release, served on reserved capacity.

Sales-assisted Early Access

How you can run it

On Token Factory

Serverless
Not available
Dedicated endpoint
Available
Modalities
Text
Context window
131K tokens

Run this model on a dedicated endpoint today, or talk to us about serverless access.

The model

Developed by
OpenAI
Family
gpt-oss
Parameters
117B
Licence
Apache 2.0

Before you request access

What happens after I request access?
We scope the workload with you, confirm what we can serve, set up your account and provision access.
What does a dedicated endpoint commit me to?
Capacity is allocated in whole 8-GPU nodes and carries a minimum term. We confirm sizing and term before anything is reserved.
How are model changes and deprecations handled?
Catalog changes are communicated before they take effect. Ask us for the current policy on this model.

What the call covers

Workload fit for gpt-oss-120b, which deployment mode suits it, what we can commit to serving and on what capacity, commercial terms, and onboarding.