Qwen3.5-27B pricing & benchmarks
Released Feb 25, 2026 · served by 6 providers
What Qwen3.5-27B is
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
- Model ID:
qwen/qwen3.5-27b - Modality: text+image+video->text
- Tool calling: supported · Structured output: supported
- Weights:
Qwen/Qwen3.5-27Bon Hugging Face
Providers and prices for Qwen3.5-27B
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Alibabacheapest | $0.195 | $1.56 | 262K | — | 99.7% | — |
| SiliconFlow | $0.250 | $2.00 | 262K | fp8 | 98.3% | — |
| AtlasCloud | $0.270 | $2.16 | 262K | fp8 | 98.9% | — |
| Novita | $0.300 | $2.40 | 262K | bf16 | 99.7% | — |
| Phala | $0.300 | $2.40 | 262K | — | 97.7% | — |
| DeepInfra | $0.260 | $2.60 | 262K | fp8 | 99.1% | — |
FAQ
How much does the Qwen3.5-27B API cost?
$0.195 per 1M input tokens and $1.56 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.66.
Is Qwen3.5-27B free?
No. It is a paid model starting at $0.195 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Qwen3.5-27B?
Alibaba at $0.195 input / $1.56 output per 1M tokens.
Can I self-host Qwen3.5-27B?
Yes — weights are published as Qwen/Qwen3.5-27B on Hugging Face, so you can run it on your own hardware.