Benchmarks

Z.ai: GLM 4.6

z-ai/glm-4.6

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

Prompt $/1M
$0.500
Completion $/1M
$2.000
Context
202,752 tok
Best provider
125 tok/s · BaseTen

Specs

Model id
z-ai/glm-4.6
Display name
Z.ai: GLM 4.6
Modality
text->text
Tokenizer
Other
Instruct type
Context length
202,752
Max completion
131,072
Prompt $/1M
$0.500
Completion $/1M
$2.000
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
BaseTenfastest125
Artificial Analysis68
AtlasCloudfastest31
DeepInfra26
Z.AI26
Novita25
SiliconFlow18
Venice17
Chutes14
Io Net4
Z.ai: GLM 4.6 - Benchmarks, Pricing and Speed | TryAii