Ling-3.0-flash pricing
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers... Live index: 3 priced offers. Best input $0.021 per million tokens from Openrouter. Best output $0.063 per million tokens from Openrouter.
Pricing across providers
All figures are list prices per million tokens unless a column says otherwise. 3 offers are listed for Ling-3.0-flash. Best input in this view: Openrouter.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
D DeepInfra | $0.060 | $0.180 | $0.012 | — |
N Novita | $0.060 | $0.180 | $0.012 | — |
O Openrouter | $0.021 | $0.063 | $0.0042 | — |
Input vs output · per provider
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Ling-3.0-flash cost scales with traffic.
Provider
0.006000¢ / req
0.009000¢ / req
Model specifications
Quick spec sheet for Ling-3.0-flash before you dive back into pricing. Reported under Inclusionai.
- Context window
- 262,144 tokens
- Max output
- 32,768 tokens
- Vision (images)
- No
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Jul 2026
- Primary provider
- Inclusionai
- Model family
- N/A
Compare Ling-3.0-flash
Open a pair page to see Ling-3.0-flash next to another model with a shared provider matrix. 6 shortcuts below.
Locked
Compare with
Pick a model on both sides.
Popular Ling-3.0-flash comparisons
- Ling-3.0-flash vs GPT-4o
Compare pricing side by side
- Ling-3.0-flash vs Claude Sonnet 4.6
Compare pricing side by side
- Ling-3.0-flash vs Gemini 2.0 Flash
Compare pricing side by side
- Ling-3.0-flash vs Llama 3.1 70B
Compare pricing side by side
- Ling-3.0-flash vs Mistral 7B
Compare pricing side by side
- Ling-3.0-flash vs GLM 4.7
Compare pricing side by side
Frequently asked questions
Quick frequently asked items for Ling-3.0-flash pricing and limits. The short model note from our index: *Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key...
Related models
Models close to Ling-3.0-flash by family, provider, and context window — each with live pricing in our catalog.