GLM 5.1
ZhipuGLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
In
$0.97
Out
$3.04
Blended
$1.48
Where to buy this model
14 access routes| Route | In | Out | Blended | vs |
|---|---|---|---|---|
| ● StreamLake cheapestfp8 | $0.97 | $3.04 | $1.48 | - |
| ● Chutes fp8 | $0.98 | $3.08 | $1.51 | +2% |
| $1.05 | $3.50 | $1.66 | +12% | |
| ● SiliconFlow fp8 | $1.19 | $3.74 | $1.83 | +24% |
| ● AtlasCloud fp8 | $1.26 | $3.96 | $1.94 | +31% |
| ● Phala | $1.21 | $4.20 | $1.96 | +32% |
| ● Alibaba fp8 | $1.33 | $4.18 | $2.04 | +38% |
| $1.38 | $4.40 | $2.13 | +44% | |
| $1.40 | $4.40 | $2.15 | +45% | |
| ● Baidu fp8 | $1.40 | $4.40 | $2.15 | +45% |
| $1.40 | $4.40 | $2.15 | +45% | |
| ● Friendli | $1.40 | $4.40 | $2.15 | +45% |
| ● Z.AI fp8 | $1.40 | $4.40 | $2.15 | +45% |
| ● Venice fp8 | $1.40 | $4.40 | $2.15 | +45% |
● official API · ● cloud platform · ● inference host · USD per 1M tokens. Same model, different providers - price and quantization vary.
- Ctx
- 205K
- max output
- 128K
- cache read /1M
- $0.18
- Family
- glm
- Modalities
- text
- Features
- json · reasoning · tools
Live pricing, updated 2026-09-15.