Qwen3 235B A22B Instruct 2507
AlibabaQwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
In
$0.09
Out
$0.35
Blended
$0.15
Where to buy this model
10 access routes| Route | In | Out | Blended | vs |
|---|---|---|---|---|
| $0.09 | $0.35 | $0.15 | - | |
| $0.09 | $0.55 | $0.21 | +40% | |
| $0.09 | $0.58 | $0.21 | +40% | |
| ● Alibaba | $0.15 | $0.60 | $0.26 | +73% |
| ● Venice fp8 | $0.15 | $0.75 | $0.30 | +100% |
| $0.20 | $0.60 | $0.30 | +100% | |
| $0.14 | $0.80 | $0.31 | +107% | |
| ● StreamLake | $0.21 | $0.84 | $0.37 | +147% |
| ● AtlasCloud fp8 | $0.20 | $0.88 | $0.37 | +147% |
| $0.22 | $0.88 | $0.39 | +160% |
● official API · ● cloud platform · ● inference host · USD per 1M tokens. Same model, different providers - price and quantization vary.
- Ctx
- 262K
- max output
- 236K
- cache read /1M
- $0.02
- Family
- qwen
- Modalities
- text
- Features
- json · tools
Live pricing, updated 2026-09-15.