Mercury 2.5 pricing
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving... Below you will find 2 current rows with input and output dollars per million. Right now the lowest input is $0.040 and the lowest output is $0.150.
Pricing across providers
All figures are list prices per million tokens unless a column says otherwise. 2 offers are listed for Mercury 2.5. Best input in this view: Openrouter.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
I Inception | $0.200 | $0.750 | $0.020 | — |
O Openrouter | $0.040 | $0.150 | — | — |
Input vs output · per provider
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Mercury 2.5 cost scales with traffic.
Provider
0.020000¢ / req
0.037500¢ / req
Model specifications
Context length, caps, and capability flags for Mercury 2.5. Values follow the main provider (Inception) record in our index.
- Context window
- 260,000 tokens
- Max output
- 65,536 tokens
- Vision (images)
- No
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Sep 2026
- Primary provider
- Inception
- Model family
- N/A
Compare Mercury 2.5
These links open full side by side pages for Mercury 2.5. We picked pairs that people often shop together. 6 ready to open.
Locked
Compare with
Pick a model on both sides.
Popular Mercury 2.5 comparisons
- Mercury 2.5 vs GPT-4o
Compare pricing side by side
- Mercury 2.5 vs Claude Sonnet 4.6
Compare pricing side by side
- Mercury 2.5 vs Gemini 2.0 Flash
Compare pricing side by side
- Mercury 2.5 vs Llama 3.1 70B
Compare pricing side by side
- Mercury 2.5 vs Mistral 7B
Compare pricing side by side
- Mercury 2.5 vs GLM 4.7
Compare pricing side by side
Frequently asked questions
Answers pull from the same numbers you see on this page. The short model note from our index: Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Related models
Models close to Mercury 2.5 by family, provider, and context window — each with live pricing in our catalog.