Ling-3.0-flash (free) pricing & benchmarks
Released Jul 23, 2026 · served by 1 provider
What Ling-3.0-flash (free) is
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
- Model ID:
inclusionai/ling-3.0-flash:free - Modality: text->text
- Tool calling: supported · Structured output: not supported
Providers and prices for Ling-3.0-flash (free)
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Novitacheapest | Free | Free | 262K | — | 100.0% | — |
FAQ
How much does the Ling-3.0-flash (free) API cost?
Free per 1M input tokens and Free per 1M output tokens.
Is Ling-3.0-flash (free) free?
Yes — it is served at $0 for both input and output tokens on OpenRouter, subject to rate limits.
Which provider is cheapest for Ling-3.0-flash (free)?
Novita at Free input / Free output per 1M tokens.
Can I self-host Ling-3.0-flash (free)?
No. Ling-3.0-flash (free) is closed-weight and only available through hosted APIs.