Alibaba Qwen · Reasoning
Qwen3-30B-A3B-Thinking-2507
Small MoE reasoner that self-hosts on a single 48GB card in FP8; output pricing is high relative to size, so prefer local deployment over the API for volume.
Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation
Model ID qwen3-30b-a3b-thinking-2507
- Parameters
- 30B total · 3B active30B-A3B MoE
- Architecture
- Mixture of experts
- Context window
- 262,144 tokensMaximum input
- Max output
- 32,768 tokensPer response
- Input price
- $0.20 per 1M tokens
- Output price
- $2.40 per 1M tokens
- Cached input
- Not published
- Blended price
- $0.75 per 1M tokens3:1 input to output
- Knowledge cutoff
- Not recorded
- Released
- 2025-07
- Status
- ga
- Size band
- 15B to 40B
License and openness
Qwen3-30B-A3B-Thinking-2507 is released under the Apache-2.0. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.
Capabilities
| Input modalities | text |
|---|---|
| Output modalities | text |
What it is for
At a three-to-one input-to-output ratio, Qwen3-30B-A3B-Thinking-2507 costs $0.75 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $29.76 per month at list price, before caching or batch discounts.
Cost at list price
| Input | $0.20 per 1M tokens |
|---|---|
| Output | $2.40 per 1M tokens |
| Cached input | Not published |
| Blended, 3:1 | $0.75 per 1M tokens |
Sources
- Source: Alibaba Qwen pricing ↗
- API docs ↗
- Full reference record for Qwen3-30B-A3B-Thinking-2507 ↗
- Alibaba Qwen in the reference ↗
Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.
Other open-weight models from Alibaba Qwen
- Qwen3.8-2.4T-A95B · 2.4T total · 95B active
- Qwen3.8-27B · 27B dense
- Qwen3.6-35B-A3B · 35B total · 3B active
- Qwen3.6-27B · 27B dense
- Qwen3.5-397B-A17B · 397B total · 17B active
- Qwen3.5-122B-A10B · 122B total · 10B active
- Qwen3.5-35B-A3B · 35B total · 3B active
- Qwen3.5-27B · 27B dense
- Qwen3-Coder-Next · Not recorded
- Qwen3-Coder-480B-A35B-Instruct · 480B total · 35B active
- Qwen3-Coder-30B-A3B-Instruct · 30B total · 3B active
- Qwen3-Next-80B-A3B-Thinking · 80B total · 3B active