Ling-2.6-flash pricing
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency.... Below you will find 1 current row with input and output dollars per million. Right now the lowest input is $0.010 and the lowest output is $0.030.
Pricing across providers
Every row is a seller of Ling-2.6-flash with token pricing we track. The cheapest input in this snapshot is from Openrouter. The bar chart shows the same input and output dollars per million for a quick scan.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
O Openrouter | $0.010 | $0.030 | — | — |
Input vs output · 1M tokens
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Ling-2.6-flash cost scales with traffic.
0.001000¢ / req
0.001500¢ / req
Model specifications
These fields describe Ling-2.6-flash as we store it (source: Inclusionai). They sit next to price so buyers can check limits and tools in one place.
- Context window
- 262,144 tokens
- Max output
- 32,768 tokens
- Vision (images)
- No
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Apr 2026
- Primary provider
- Inclusionai
- Model family
- N/A
Compare Ling-2.6-flash
Jump into a comparison when you want one table for two models instead of two tabs. 6 curated matches for Ling-2.6-flash.
Locked
Compare with
Pick a model on both sides.
Popular Ling-2.6-flash comparisons
- Ling-2.6-flash vs GPT-4o
Compare pricing side by side
- Ling-2.6-flash vs Claude Sonnet 4.6
Compare pricing side by side
- Ling-2.6-flash vs Gemini 2.0 Flash
Compare pricing side by side
- Ling-2.6-flash vs Llama 3.1 70B
Compare pricing side by side
- Ling-2.6-flash vs Mistral 7B
Compare pricing side by side
- Ling-2.6-flash vs GLM 4.7
Compare pricing side by side
Frequently asked questions
Answers pull from the same numbers you see on this page. The short model note from our index: Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficienc...
Related models
Models close to Ling-2.6-flash by family, provider, and context window — each with live pricing in our catalog.