Qwen2.5-Omni-7B
The small open any-to-any model people actually self-host for offline voice assistants; output pricing varies by modality and is not stated as a single figure on the pricing page.
Data checked 2026-09-06
Alibaba Qwen
ga
open weights
qwen2.5-omni-7b
- Context window
- 32.8K tokens in
- Max output
- 2K tokens out
- Input
- $0.10 per 1M tokens
- Output
- — per 1M tokens
- Released
- 2025-03
- Parameters
- 7B dense
- Licence
- Apache-2.0
A What it is for
Superseded by the Qwen3.5-Omni hosted tier.
B Capabilities
vision streaming fine-tuning
| Input modalities | text, image, audio, video |
|---|---|
| Output modalities | text, audio |
C Others from Alibaba Qwen
| Model | Context | Max out | In $/M | Out $/M | Blended | Status |
|---|---|---|---|---|---|---|
Qwen3.8-Max
qwen3.8-max
|
1M | 131.1K | $2.00 | $6.00 | $3.00 | ga |
Qwen3.7-Max
qwen3.7-max
|
1M | 131.1K | $2.50 | $7.50 | $3.75 | ga |
Qwen3-Max
qwen3-max
|
262.1K | 65.5K | $1.20 | $6.00 | $2.40 | ga |
Qwen-Max (legacy)
qwen-max
|
32.8K | 8.2K | $1.60 | $6.40 | $2.80 | deprecated |
Qwen3.7-Plus
qwen3.7-plus
|
1M | 131.1K | $0.40 | $1.60 | $0.70 | ga |
Qwen3.5-Plus
qwen3.5-plus
|
1M | 131.1K | $0.40 | $2.40 | $0.90 | ga |
Qwen-Plus (rolling)
qwen-plus
|
1M | 32.8K | $0.40 | $1.20 | $0.60 | ga |
Qwen3.8-Flash
qwen3.8-flash
|
1M | 131.1K | $0.15 | $0.47 | $0.23 | ga |
Qwen3.7-Flash
qwen3.7-flash
|
1M | 65.5K | $0.03 | $0.13 | $0.06 | ga |
Qwen3.5-Flash
qwen3.5-flash
|
1M | 65.5K | $0.10 | $0.40 | $0.18 | ga |
Qwen-Flash (rolling)
qwen-flash
|
1M | 32.8K | $0.05 | $0.40 | $0.14 | ga |
Qwen-Turbo
qwen-turbo
|
1M | 16.4K | $0.05 | $0.20 | $0.09 | deprecated |
D Verify before you commit
Model pricing changes without notice and this page is a snapshot. Confirm against the vendor's own page before you build a budget on it.
Source: alibabacloud.com · checked 2026-09-06 · Alibaba Qwen pricing · API docs
Some figures on this page are recorded at medium confidence — the vendor does not publish them in a single authoritative place. Treat them as indicative.