← All open-weight models

Alibaba Qwen · Vision & multimodal

Qwen3.6-35B-A3B

Only 3B active parameters, so it runs fast on modest hardware while punching well above its weight on coding benchmarks — the sweet spot for local agentic coding under Apache 2.0. Thinking is enabled by default on the Qwen3.6 series.

Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation

Model ID qwen3.6-35b-a3b

Parameters
35B total · 3B active35B-A3B MoE
Architecture
Mixture of experts
Context window
262,144 tokensMaximum input
Max output
65,536 tokensPer response
Input price
$0.375 per 1M tokens
Output price
$2.25 per 1M tokens
Cached input
Not published
Blended price
$0.8438 per 1M tokens3:1 input to output
Knowledge cutoff
Not recorded
Released
2026-04-16
Status
ga
Size band
15B to 40B

License and openness

Qwen3.6-35B-A3B is released under the Apache-2.0. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.

Capabilities

Input modalitiestext, image
Output modalitiestext

What it is for

Thinking is enabled by default on the Qwen3.6 series.

At a three-to-one input-to-output ratio, Qwen3.6-35B-A3B costs $0.84 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $33.52 per month at list price, before caching or batch discounts.

Cost at list price

Input$0.375 per 1M tokens
Output$2.25 per 1M tokens
Cached inputNot published
Blended, 3:1$0.8438 per 1M tokens
Estimate a workload across all models ↗

Sources

Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.

Other open-weight models from Alibaba Qwen

See all 31 from Alibaba Qwen ↗

Browse all open-weight models ↗