Mercury 2.5 Preview pricing
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving... Below you will find 1 current row with input and output dollars per million. Right now the lowest input is $0.040 and the lowest output is $0.150.
Pricing across providers
Use this table to read Mercury 2.5 Preview list prices. We show 1 source right now. Lowest input in the grid: Openrouter. The chart below the table helps when output prices are much higher than input prices.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
O Openrouter | $0.040 | $0.150 | — | — |
Input vs output · 1M tokens
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Mercury 2.5 Preview cost scales with traffic.
0.004000¢ / req
0.007500¢ / req
Model specifications
These fields describe Mercury 2.5 Preview as we store it (source: Inception). They sit next to price so buyers can check limits and tools in one place.
- Context window
- 260,000 tokens
- Max output
- 65,536 tokens
- Vision (images)
- No
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Aug 2026
- Primary provider
- Inception
- Model family
- N/A
Compare Mercury 2.5 Preview
Open a pair page to see Mercury 2.5 Preview next to another model with a shared provider matrix. 6 shortcuts below.
Locked
Compare with
Pick a model on both sides.
Popular Mercury 2.5 Preview comparisons
- Mercury 2.5 Preview vs GPT-4o
Compare pricing side by side
- Mercury 2.5 Preview vs Claude Sonnet 4.6
Compare pricing side by side
- Mercury 2.5 Preview vs Gemini 2.0 Flash
Compare pricing side by side
- Mercury 2.5 Preview vs Llama 3.1 70B
Compare pricing side by side
- Mercury 2.5 Preview vs Mistral 7B
Compare pricing side by side
- Mercury 2.5 Preview vs GLM 4.7
Compare pricing side by side
Frequently asked questions
Quick frequently asked items for Mercury 2.5 Preview pricing and limits. The short model note from our index: Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Related models
Models close to Mercury 2.5 Preview by family, provider, and context window — each with live pricing in our catalog.