inclusionai Reasoning Free tier

Ling-3.0-flash (free) pricing & benchmarks

Released Jul 23, 2026 · served by 1 provider

Input / 1M tokensFree
Output / 1M tokensFree
Context window262K
Cached input
1M in + 300K out$0.00

What Ling-3.0-flash (free) is

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Providers and prices for Ling-3.0-flash (free)

The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.

ProviderInputOutputContext QuantUptime 24hThroughput
Novitacheapest Free Free 262K 100.0%

FAQ

How much does the Ling-3.0-flash (free) API cost?

Free per 1M input tokens and Free per 1M output tokens.

Is Ling-3.0-flash (free) free?

Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.

Which provider is cheapest for Ling-3.0-flash (free)?

Novita at Free input / Free output per 1M tokens.

Can I self-host Ling-3.0-flash (free)?

No. Ling-3.0-flash (free) is closed-weight and only available through hosted APIs.