Qwen3 VL 235B API Pricing
Qwen3 VL 235B by Alibaba costs $0.40 per million input tokens and $1.60 per million output tokens.
Price verified from Alibaba's pricing page.
Rates and limits
- Input / 1M
- $0.40
- Output / 1M
- $1.60
- Cached input / 1M
- Not listed
- Batch discount
- None
- Context window
- 262K tokens
- Released
- October 2025
Pricing note: Price is the instruct (non-thinking) variant, qwen3-vl-235b-a22b-instruct, on Alibaba Model Studio International (Singapore). The thinking variant, qwen3-vl-235b-a22b-thinking, is $0.4 input / $4 output. Global and China (Beijing) deployments are priced differently ($0.287 / $1.147 for instruct).
What Qwen3 VL 235B costs per month
Token volume split 3:1 between input and output, with no prompt caching.
| Tokens per month | Input / output | Standard |
|---|---|---|
| 1M | 750K / 250K | $0.70 |
| 10M | 7.5M / 2.5M | $7.00 |
| 100M | 75M / 25M | $70.00 |
Where to run Qwen3 VL 235B
Prices above are Alibaba's direct rates. Cloud platforms can charge differently; the cloud platform view compares them.
- Alibaba DashScope (Model Studio)Direct API
- OpenRouterCloud / platform
Qwen3 VL 235B pricing FAQ
How much does Qwen3 VL 235B cost per million tokens?
Qwen3 VL 235B costs $0.40 per million input tokens and $1.60 per million output tokens, as listed on Alibaba's pricing page on September 30, 2026.
How much does Qwen3 VL 235B cost for 10 million tokens a month?
About $7.00 a month, assuming 7.5 million input and 2.5 million output tokens with no prompt caching.
What is the context window of Qwen3 VL 235B?
Qwen3 VL 235B accepts up to 262,144 tokens of context.
Does Qwen3 VL 235B support prompt caching?
No cached-input price is listed for Qwen3 VL 235B, so repeated prompt prefixes are billed at the standard input rate.
Is there a batch discount for Qwen3 VL 235B?
No batch API price is listed for Qwen3 VL 235B.
Where can I use the Qwen3 VL 235B API?
Qwen3 VL 235B is available through Alibaba DashScope (Model Studio) and OpenRouter.
Compare with other models
More from Alibaba
Prices shown as input / output per million tokens. All model prices · Side-by-side comparison