Z.ai: GLM 4.6
z-ai/glm-4.6
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
Prompt $/1M
$0.500
Completion $/1M
$2.000
Context
202,752 tok
Best provider
125 tok/s · BaseTen
Specs
- Model id
- z-ai/glm-4.6
- Display name
- Z.ai: GLM 4.6
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 202,752
- Max completion
- 131,072
- Prompt $/1M
- $0.500
- Completion $/1M
- $2.000
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| BaseTenfastest | 125 |
| Artificial Analysis | 68 |
| AtlasCloudfastest | 31 |
| DeepInfra | 26 |
| Z.AI | 26 |
| Novita | 25 |
| SiliconFlow | 18 |
| Venice | 17 |
| Chutes | 14 |
| Io Net | 4 |