qwen Open weights

Qwen3 Coder Next pricing & benchmarks

Released Feb 4, 2026 · served by 5 providers

Input / 1M tokens$0.110
Output / 1M tokens$0.800
Context window262K
Cached input$0.070
Intelligence index21.1
Coding index36.2
Agentic index8.8
1M in + 300K out$0.35

What Qwen3 Coder Next is

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

Providers and prices for Qwen3 Coder Next

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Ionstreamcheapest $0.110 $0.800 262K fp8 97.2%
Parasail $0.120 $0.800 262K bf16 100.0%
StreamLake $0.180 $0.900 256K 99.5%
Novita $0.200 $1.50 262K fp8 99.4%
Alibaba $0.300 $1.50 262K 99.5%

Cheaper models in the same class

Models scoring within 4 points of Qwen3 Coder Next on the Intelligence Index, but with a lower output price.

FAQ

How much does the Qwen3 Coder Next API cost?

$0.110 per 1M input tokens and $0.800 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.35.

Is Qwen3 Coder Next free?

No. It is a paid model starting at $0.110 per 1M input tokens, though some providers offer trial credits.

Which provider is cheapest for Qwen3 Coder Next?

Ionstream at $0.110 input / $0.800 output per 1M tokens (fp8 quantization).

Can I self-host Qwen3 Coder Next?

Yes — weights are published as Qwen/Qwen3-Coder-Next on Hugging Face, so you can run it on your own hardware.