OpenAI: gpt-oss-120b
openai/gpt-oss-120b
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
Prompt $/1M
$0.037
Completion $/1M
$0.170
Context
131,072 tok
Best provider
573 tok/s · Cerebras
Specs
- Model id
- openai/gpt-oss-120b
- Display name
- OpenAI: gpt-oss-120b
- Modality
- text->text
- Tokenizer
- GPT
- Instruct type
- —
- Context length
- 131,072
- Max completion
- 131,072
- Prompt $/1M
- $0.037
- Completion $/1M
- $0.170
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Cerebrasfastest | 573 |
| Artificial Analysis | 343 |
| Groq | 331 |
| Amazon Bedrock | 283 |
| SambaNova | 231 |
| Nebius | 153 |
| DeepInfra | 152 |
| Fireworks | 126 |
| Mara | 117 |
| AtlasCloud | 112 |
| Parasail | 86 |
| Ambient | 72 |
| Together | 70 |
| Novita | 57 |
| Io Net | 53 |
| Phala | 46 |
| 44 | |
| Clarifai | 40 |
| OpenInference | 36 |
| BaseTen | 31 |
| WandB | 25 |
| Mancer 2 | 24 |
| DekaLLM | 21 |
| DigitalOcean | 18 |
| SiliconFlow | 10 |
| Chutes | 5 |