All models
GLM 5.2
z-ai/glm-5.2
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Modalities
Text
In / Out price
$0.63 / $1.98 per 1M
Context
1.05M
Released
Jun 13, 2026
Providers
Live list prices and uptime across 31 providers serving this model.
| Provider | Input /M | Output /M | Cache read /M | Context | Max output | Uptime 24h |
|---|---|---|---|---|---|---|
| Novitafp8 | $0.461 | $1.448 | $0.086 | 1.05M | 131K | 98.08% |
| Baidufp8 | $0.462 | $1.452 | $0.086 | 1.05M | 131K | 95.39% |
| Sail Researchfp8 | $0.50 | $3.15 | $0.115 | 1.05M | 131K | 99.56% |
| StreamLakefp8 | $0.56 | $1.76 | $0.104 | 1.02M | 128K | 99.23% |
| DigitalOcean | $0.70 | $2.20 | $0.105 | 262K | — | 98.50% |
| GMICloudfp8 | $0.742 | $2.332 | $0.138 | 1.05M | — | 97.58% |
| DeepInfrafp4 | $0.75 | $2.40 | $0.14 | 1.05M | 164K | 98.86% |
| Inceptronfp4 | $0.75 | $2.90 | $0.17 | 1.05M | 1.05M | 97.23% |
| CoreWeavefp4 | $0.76 | $2.42 | $0.14 | 262K | 262K | 84.22% |
| Decartfp4 | $0.768 | $2.56 | $0.128 | 1.05M | 1.05M | 98.49% |
| AkashMLfp8 | $0.77 | $2.42 | $0.143 | 97K | 97K | 98.95% |
| Alibabafp8 | $0.966 | $3.036 | $0.193 | 1.05M | 131K | 99.75% |
| Ambientfp8 | $1.05 | $4.40 | $0.20 | 203K | 203K | 92.41% |
| Morphfp4 | $1.10 | $4.10 | $0.22 | 1.05M | 1.05M | 99.94% |
| Phalafp8 | $1.134 | $3.00 | $0.211 | 1.05M | 131K | 96.46% |
| SiliconFlowfp8 | $1.19 | $3.74 | $0.221 | 1.05M | 262K | 86.85% |
| Waferfp4 | $1.26 | $3.96 | $0.234 | 1.05M | 131K | 95.34% |
| AtlasCloudfp8 | $1.26 | $3.96 | $0.234 | 1.05M | 131K | 99.31% |
| Z.AIfp8 | $1.40 | $4.40 | $0.26 | 1.05M | 131K | 99.24% |
| Fireworks | $1.40 | $4.40 | $0.14 | 1.05M | — | 98.96% |
| Cloudflare | $1.40 | $4.40 | $0.26 | 262K | 262K | 99.94% |
| Friendli | $1.40 | $4.40 | $0.26 | 1.05M | 1.05M | 99.36% |
| Parasailfp4 | $1.40 | $4.40 | $0.26 | 262K | 262K | 99.85% |
| Venicefp8 | $1.40 | $4.40 | $0.26 | 1M | 131K | 96.74% |
| Together | $1.40 | $4.40 | $0.26 | 512K | — | 94.17% |
| Crusoefp8 | $1.40 | $4.40 | $0.26 | 1.05M | — | 99.22% |
| BaseTenfp8 | $1.40 | $4.40 | $0.14 | 1.05M | 262K | 99.77% |
| Fireworks | $2.10 | $6.60 | $0.21 | 1.05M | — | 95.96% |
| Cloudflare | $2.10 | $6.60 | $0.21 | 262K | 262K | 99.85% |
| BaseTenfp8 | $2.10 | $6.60 | $0.21 | 1.05M | 262K | 99.65% |
| Alibabafp8 | $2.31 | $7.26 | $0.462 | 1.05M | 131K | 99.81% |
No implicit prompt caching; cache reads require explicit cache control.
Capabilities
Tool calling
Reasoning
frequency_penalty
include_reasoning
logit_bias
logprobs
max_tokens
min_p
parallel_tool_calls
presence_penalty
reasoning
reasoning_effort
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
top_k
top_logprobs
top_p
Pricing detail
- Input
- $0.63
- Output
- $1.98
- Cache read
- $0.095
- Cache write
- —
USD per 1M tokens.