Ling-2.6-flash pricing & benchmarks
Released Apr 21, 2026 · served by 1 provider
What Ling-2.6-flash is
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
- Model ID:
inclusionai/ling-2.6-flash - Modality: text->text
- Tool calling: supported · Structured output: supported
Providers and prices for Ling-2.6-flash
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Novitacheapest | $0.010 | $0.030 | 262K | — | 100.0% | — |
Cheaper models in the same class
Models scoring within 4 points of Ling-2.6-flash on the Intelligence Index, but with a lower output price.
FAQ
How much does the Ling-2.6-flash API cost?
$0.010 per 1M input tokens and $0.030 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.02.
Is Ling-2.6-flash free?
No. It is a paid model starting at $0.010 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Ling-2.6-flash?
Novita at $0.010 input / $0.030 output per 1M tokens.
Can I self-host Ling-2.6-flash?
No. Ling-2.6-flash is closed-weight and only available through hosted APIs.