qwen Open weights Vision

Qwen3 VL 235B A22B Instruct pricing & benchmarks

Released Sep 23, 2025 · served by 5 providers

Input / 1M tokens$0.200
Output / 1M tokens$0.880
Context window262K
Cached input$0.100
1M in + 300K out$0.46

What Qwen3 VL 235B A22B Instruct is

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...

Providers and prices for Qwen3 VL 235B A22B Instruct

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
DeepInfracheapest $0.200 $0.880 262K fp8 88.5%
Alibaba $0.260 $1.04 131K 99.6%
Novita $0.300 $1.50 131K bf16 96.8%
Venice $0.210 $1.90 128K fp8 97.9%
Parasail $0.210 $1.90 131K fp8 99.4%

FAQ

How much does the Qwen3 VL 235B A22B Instruct API cost?

$0.200 per 1M input tokens and $0.880 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.46.

Is Qwen3 VL 235B A22B Instruct free?

No. It is a paid model starting at $0.200 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for Qwen3 VL 235B A22B Instruct?

DeepInfra at $0.200 input / $0.880 output per 1M tokens (fp8 quantization).

Can I self-host Qwen3 VL 235B A22B Instruct?

Yes — weights are published as Qwen/Qwen3-VL-235B-A22B-Instruct on Hugging Face, so you can run it on your own hardware.