OpenAI
gpt-oss-120b
OpenAI's open-weight release, served on reserved capacity.
Sales-assisted Early Access
How you can run it
On Token Factory
- Serverless
- Not available
- Dedicated endpoint
- Available
- Modalities
- Text
- Context window
- 131K tokens
Run this model on a dedicated endpoint today, or talk to us about serverless access.
The model
- Developed by
- OpenAI
- Family
- gpt-oss
- Parameters
- 117B
- Licence
- Apache 2.0
Before you request access
- What happens after I request access?
- We scope the workload with you, confirm what we can serve, set up your account and provision access.
- What does a dedicated endpoint commit me to?
- Capacity is allocated in whole 8-GPU nodes and carries a minimum term. We confirm sizing and term before anything is reserved.
- How are model changes and deprecations handled?
- Catalog changes are communicated before they take effect. Ask us for the current policy on this model.
What the call covers
Workload fit for gpt-oss-120b, which deployment mode suits it, what we can commit to serving and on what capacity, commercial terms, and onboarding.