Z.ai: GLM 4.7 Flash
z-ai/glm-4.7-flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
Prompt $/1M
$0.060
Completion $/1M
$0.400
Context
200,000 tok
Best provider
128 tok/s · Artificial Analysis
Specs
- Model id
- z-ai/glm-4.7-flash
- Display name
- Z.ai: GLM 4.7 Flash
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 200,000
- Max completion
- 131,072
- Prompt $/1M
- $0.060
- Completion $/1M
- $0.400
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 128 |
| Phala | 43 |
| Novitafastest | 37 |
| Z.AI | 30 |
| DeepInfra | 24 |
| Cloudflare | 19 |
| Venice | 12 |