Qwen3.5 397B A17B
AlibabaThe Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
In
$0.55
Out
$3.50
Blended
$1.29
Where to buy this model
10 access routes| Route | In | Out | Blended | vs |
|---|---|---|---|---|
| ● Alibaba cheapest | $0.39 | $2.34 | $0.88 | - |
| $0.45 | $3.00 | $1.09 | +24% | |
| $0.50 | $3.60 | $1.27 | +44% | |
| $0.55 | $3.50 | $1.29 | +47% | |
| ● Phala | $0.55 | $3.50 | $1.29 | +47% |
| ● AtlasCloud fp8 | $0.55 | $3.50 | $1.29 | +47% |
| ● StreamLake | $0.60 | $3.60 | $1.35 | +53% |
| $0.60 | $3.60 | $1.35 | +53% | |
| $0.60 | $3.60 | $1.35 | +53% | |
| ● Venice | $0.75 | $4.50 | $1.69 | +92% |
● official API · ● cloud platform · ● inference host · USD per 1M tokens. Same model, different providers - price and quantization vary.
- Ctx
- 262K
- max output
- 236K
- cache read /1M
- $0.23
- Family
- qwen
- Modalities
- text · image · video
- Features
- json · reasoning · tools
Live pricing, updated 2026-09-15.