GLM 4.6
ZhipuCompared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
In
$0.43
Out
$1.75
Blended
$0.76
Where to buy this model
5 access routes| Route | In | Out | Blended | vs |
|---|---|---|---|---|
| ● Venice cheapestfp4 | $0.43 | $1.75 | $0.76 | - |
| $0.50 | $2.00 | $0.88 | +16% | |
| $0.55 | $2.20 | $0.96 | +26% | |
| ● AtlasCloud fp8 | $0.60 | $2.20 | $1.00 | +32% |
| ● Z.AI fp4 | $0.60 | $2.20 | $1.00 | +32% |
● official API · ● cloud platform · ● inference host · USD per 1M tokens. Same model, different providers - price and quantization vary.
- Ctx
- 205K
- max output
- 16K
- cache read /1M
- $0.08
- Family
- glm
- Modalities
- text
- Features
- json · reasoning · tools
Live pricing, updated 2026-09-15.