Ling-3.0-flash pricing
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers... Live index: 1 priced offer. Best input $0.021 per million tokens from Openrouter. Best output $0.063 per million tokens from Openrouter.
Pricing across providers
All figures are list prices per million tokens unless a column says otherwise. 1 offer is listed for Ling-3.0-flash. Best input in this view: Openrouter.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
O Openrouter | $0.021 | $0.063 | — | — |
Input vs output · 1M tokens
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Ling-3.0-flash cost scales with traffic.
0.002100¢ / req
0.003150¢ / req
Model specifications
Quick spec sheet for Ling-3.0-flash before you dive back into pricing. Reported under Inclusionai.
- Context window
- 262,144 tokens
- Max output
- 32,768 tokens
- Vision (images)
- No
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Jul 2026
- Primary provider
- Inclusionai
- Model family
- N/A
Compare Ling-3.0-flash
Open a pair page to see Ling-3.0-flash next to another model with a shared provider matrix. 6 shortcuts below.
Locked
Compare with
Pick a model on both sides.
Popular Ling-3.0-flash comparisons
- Ling-3.0-flash vs GPT-4o
Compare pricing side by side
- Ling-3.0-flash vs GPT-4o mini
Compare pricing side by side
- Ling-3.0-flash vs Claude Sonnet 4.6
Compare pricing side by side
- Ling-3.0-flash vs Gemini 2.0 Flash
Compare pricing side by side
- Ling-3.0-flash vs o3
Compare pricing side by side
- Ling-3.0-flash vs Llama 3.1 70B
Compare pricing side by side
Frequently asked questions
Quick frequently asked items for Ling-3.0-flash pricing and limits. The short model note from our index: *Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key...
Also from Inclusionai
Other models by Inclusionai with live pricing in our catalog.