OpenAI: gpt-oss-120b
openai/gpt-oss-120b
gpt-oss-120b is a 120B open-weight model from OpenAI, ranking #127 of 235 on the TryAii Score. It's strongest on math (#16 on AIME-2025). API pricing starts at $0.037/$0.17 per million input/output tokens — about a quarter of what models at its level usually charge, with 24 independent hosts serving it. It's among the fastest models we track, at roughly 39-143 tokens/sec.
Prompt $/1M
$0.037
Completion $/1M
$0.170
Context
131,072 tok
Best provider
776 tok/s · Cerebras
Specs
- Model id
- openai/gpt-oss-120b
- Display name
- OpenAI: gpt-oss-120b
- Modality
- text->text
- Tokenizer
- GPT
- Instruct type
- —
- Context length
- 131,072
- Max completion
- 117,964
- Prompt $/1M
- $0.037
- Completion $/1M
- $0.170
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Cerebrasfastest | 776 |
| SambaNova | 315 |
| Groq | 303 |
| Amazon Bedrock | 261 |
| Nebius | 234 |
| Artificial Analysis | 159 |
| BaseTen | 157 |
| Mara | 138 |
| Fireworks | 126 |
| AtlasCloud | 112 |
| DeepInfra | 109 |
| Parasail | 103 |
| Phala | 72 |
| Ambient | 72 |
| WandB | 67 |
| Novita | 67 |
| 55 | |
| Io Net | 53 |
| Together | 41 |
| Clarifai | 40 |
| Mancer 2 | 40 |
| AkashML | 37 |
| OpenInference | 36 |
| CoreWeave | 36 |
| DigitalOcean | 28 |
| DekaLLM | 23 |
| SiliconFlow | 11 |
| Chutes | 5 |
| openaifastest | — |