Qwen3.5-Flash pricing & benchmarks
Released Feb 25, 2026 · served by 1 provider
What Qwen3.5-Flash is
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
- Model ID:
qwen/qwen3.5-flash-02-23 - Modality: text+image+video->text
- Tool calling: supported · Structured output: supported
Providers and prices for Qwen3.5-Flash
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Alibabacheapest | $0.065 | $0.260 | 1M | — | 100.0% | — |
FAQ
How much does the Qwen3.5-Flash API cost?
$0.065 per 1M input tokens and $0.260 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.14.
Is Qwen3.5-Flash free?
No. It is a paid model starting at $0.065 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Qwen3.5-Flash?
Alibaba at $0.065 input / $0.260 output per 1M tokens.
Can I self-host Qwen3.5-Flash?
No. Qwen3.5-Flash is closed-weight and only available through hosted APIs.