InclusionaiTool use

Ling-3.0-flash pricing

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers... Live index: 3 priced offers. Best input $0.021 per million tokens from Openrouter. Best output $0.063 per million tokens from Openrouter.

262K context·3 providers·verified Oct 1, 2026
Best input$0.021per 1M tokens · Openrouter
Best output$0.063per 1M tokens · Openrouter

Pricing across providers

All figures are list prices per million tokens unless a column says otherwise. 3 offers are listed for Ling-3.0-flash. Best input in this view: Openrouter.

D
DeepInfra
Input / 1M
$0.060
Output / 1M
$0.180
Cached in: $0.012
N
Novita
Input / 1M
$0.060
Output / 1M
$0.180
Cached in: $0.012
O
Openrouter
Input / 1M
$0.021
Output / 1M
$0.063
Cached in: $0.0042

Input vs output · per provider

Cost calculator

The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Ling-3.0-flash cost scales with traffic.

Provider

In: $0.060/M·Out: $0.180/M·Cache: $0.012/M

0.006000¢ / req

0.009000¢ / req

Daily
$1.50
Monthly
$45
Annual
$548

Model specifications

Quick spec sheet for Ling-3.0-flash before you dive back into pricing. Reported under Inclusionai.

Context window
262,144 tokens
Max output
32,768 tokens
Vision (images)
No
Tool / function calling
Yes
Streaming
Yes
Released
Jul 2026
Primary provider
Inclusionai
Model family
N/A

Compare Ling-3.0-flash

Open a pair page to see Ling-3.0-flash next to another model with a shared provider matrix. 6 shortcuts below.

Locked

Compare with

Pick a model on both sides.

Popular Ling-3.0-flash comparisons

Frequently asked questions

Quick frequently asked items for Ling-3.0-flash pricing and limits. The short model note from our index: *Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key...

Yes. Ling-3.0-flash is available on DeepInfra, Novita, Openrouter.

Related models

Models close to Ling-3.0-flash by family, provider, and context window — each with live pricing in our catalog.