Alibaba Qwen · General language
Qwen3.5-397B-A17B
The largest fully Apache-2.0 Qwen model — the one to self-host when licence purity matters more than raw scale and the custom-licensed Qwen3.8-2.4T is off the table. 262K native, YaRN-extensible to ~1.01M.
Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation
Model ID qwen3.5-397b-a17b
- Parameters
- 397B total · 17B active397B-A17B MoE
- Architecture
- Mixture of experts
- Context window
- 262,144 tokensMaximum input
- Max output
- 65,536 tokensPer response
- Input price
- $0.60 per 1M tokens
- Output price
- $3.60 per 1M tokens
- Cached input
- Not published
- Blended price
- $1.35 per 1M tokens3:1 input to output
- Knowledge cutoff
- 2026-01
- Released
- 2026-02-16
- Status
- ga
- Size band
- 120B and above
License and openness
Qwen3.5-397B-A17B is released under the Apache-2.0. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.
Capabilities
| Input modalities | text |
|---|---|
| Output modalities | text |
What it is for
262K native, YaRN-extensible to ~1.01M.
At a three-to-one input-to-output ratio, Qwen3.5-397B-A17B costs $1.35 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $53.64 per month at list price, before caching or batch discounts.
Cost at list price
| Input | $0.60 per 1M tokens |
|---|---|
| Output | $3.60 per 1M tokens |
| Cached input | Not published |
| Blended, 3:1 | $1.35 per 1M tokens |
Sources
- Source: Alibaba Qwen pricing ↗
- API docs ↗
- Full reference record for Qwen3.5-397B-A17B ↗
- Alibaba Qwen in the reference ↗
Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.
Other open-weight models from Alibaba Qwen
- Qwen3.8-2.4T-A95B · 2.4T total · 95B active
- Qwen3.8-27B · 27B dense
- Qwen3.6-35B-A3B · 35B total · 3B active
- Qwen3.6-27B · 27B dense
- Qwen3.5-122B-A10B · 122B total · 10B active
- Qwen3.5-35B-A3B · 35B total · 3B active
- Qwen3.5-27B · 27B dense
- Qwen3-Coder-Next · Not recorded
- Qwen3-Coder-480B-A35B-Instruct · 480B total · 35B active
- Qwen3-Coder-30B-A3B-Instruct · 30B total · 3B active
- Qwen3-Next-80B-A3B-Thinking · 80B total · 3B active
- Qwen3-Next-80B-A3B-Instruct · 80B total · 3B active