Z.ai: GLM 5.2
z-ai/glm-5.2
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Prompt $/1M
$0.818
Completion $/1M
$2.570
Context
1,048,576 tok
Best provider
198 tok/s · Artificial Analysis
Specs
- Model id
- z-ai/glm-5.2
- Display name
- Z.ai: GLM 5.2
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 1,048,576
- Max completion
- 131,072
- Prompt $/1M
- $0.818
- Completion $/1M
- $2.570
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 198 |
| Friendlifastest | 116 |
| Parasail | 103 |
| Wafer | 98 |
| Fireworks | 95 |
| WandB | 81 |
| Decart | 76 |
| Cloudflare | 75 |
| BaseTen | 66 |
| Ionstream | 65 |
| Io Net | 55 |
| Venice | 54 |
| GMICloud | 53 |
| Baidu | 50 |
| AtlasCloud | 44 |
| Morph | 43 |
| DekaLLM | 41 |
| Ambient | 40 |
| AkashML | 38 |
| Alibaba | 32 |
| StreamLake | 29 |
| Z.AI | 28 |
| DeepInfra | 25 |
| SiliconFlow | 24 |
| Phala | 20 |
| Novita | 20 |
| DigitalOcean | 16 |
| Together | 15 |
| Chutes | 14 |
| Inceptron | 4 |