Qwen3.5-9B pricing & benchmarks
Released Mar 10, 2026 · served by 5 providers
What Qwen3.5-9B is
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
- Model ID:
qwen/qwen3.5-9b - Modality: text+image+video->text
- Tool calling: supported · Structured output: supported
- Weights:
Qwen/Qwen3.5-9Bon Hugging Face
Providers and prices for Qwen3.5-9B
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| SiliconFlowcheapest | $0.100 | $0.150 | 262K | fp8 | 97.6% | — |
| DeepInfra | $0.100 | $0.150 | 262K | bf16 | 99.8% | — |
| Venice | $0.100 | $0.150 | 256K | fp8 | 99.9% | — |
| Parasail | $0.100 | $0.250 | 262K | bf16 | 94.5% | — |
| Together | $0.170 | $0.250 | 262K | — | 99.6% | — |
Cheaper models in the same class
Models scoring within 4 points of Qwen3.5-9B on the Intelligence Index, but with a lower output price.
FAQ
How much does the Qwen3.5-9B API cost?
$0.100 per 1M input tokens and $0.150 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.14.
Is Qwen3.5-9B free?
No. It is a paid model starting at $0.100 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Qwen3.5-9B?
SiliconFlow at $0.100 input / $0.150 output per 1M tokens (fp8 quantization).
Can I self-host Qwen3.5-9B?
Yes — weights are published as Qwen/Qwen3.5-9B on Hugging Face, so you can run it on your own hardware.