OpenAI: GPT-4o-mini
openai/gpt-4o-mini
GPT-4o-mini is OpenAI's budget option at $0.15/$0.6 per million input/output tokens, with a 128K-token context. It sits toward the back of the board on capability, at #208 of 235 on the TryAii Score — its best showing is #10 on GSM8K. It's among the fastest models we track, at roughly 53-76 tokens/sec.
Prompt $/1M
$0.150
Completion $/1M
$0.600
Context
128,000 tok
Best provider
88 tok/s · Azure
Specs
- Model id
- openai/gpt-4o-mini
- Display name
- OpenAI: GPT-4o-mini
- Modality
- text+image+file->text
- Tokenizer
- GPT
- Instruct type
- —
- Context length
- 128,000
- Max completion
- 16,384
- Prompt $/1M
- $0.150
- Completion $/1M
- $0.600
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedyes
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Azurefastest | 88 |
| OpenAI | 42 |
| openaifastest | — |
| Artificial Analysis | 0 |
Prompt caching by provider
| Provider | Min tokens to cache |
|---|---|
| OpenAI | 1,024 tokens |