Qwen3.8 Flash pricing
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis. Live index: 5 priced offers. Best input $0.113 per million tokens from Aihubmix. Best output $0.380 per million tokens from Aihubmix.
Pricing across providers
Every row is a seller of Qwen3.8 Flash with token pricing we track. The cheapest input in this snapshot is from Aihubmix. The bar chart shows the same input and output dollars per million for a quick scan.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
A Aihubmix | $0.113 | $0.380 | $0.014 | — |
AC Alibaba Cloud | $0.150 | $0.470 | $0.016 | — |
O Openrouter | $0.150 | $0.470 | — | — |
QA Qwen Ai Platform | $0.150 | $0.470 | $0.016 | — |
TA Together AI | $0.150 | $0.470 | — | — |
Input vs output · per provider
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how Qwen3.8 Flash cost scales with traffic.
Provider
0.011260¢ / req
0.019000¢ / req
Model specifications
Context length, caps, and capability flags for Qwen3.8 Flash. Family: Qwen. Values follow the main provider (Alibaba) record in our index.
- Context window
- 1,000,000 tokens
- Max output
- 131,072 tokens
- Vision (images)
- Yes
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Aug 2026
- Primary provider
- Alibaba
- Model family
- Qwen
Compare Qwen3.8 Flash
Open a pair page to see Qwen3.8 Flash next to another model with a shared provider matrix. 6 shortcuts below.
Locked
Compare with
Pick a model on both sides.
Popular Qwen3.8 Flash comparisons
- Qwen3.8 Flash vs GPT-4o
Compare pricing side by side
- Qwen3.8 Flash vs Claude Sonnet 4.6
Compare pricing side by side
- Qwen3.8 Flash vs Gemini 2.0 Flash
Compare pricing side by side
- Qwen3.8 Flash vs Llama 3.1 70B
Compare pricing side by side
- Qwen3.8 Flash vs Mistral 7B
Compare pricing side by side
- Qwen3.8 Flash vs GLM 4.7
Compare pricing side by side
Frequently asked questions
Answers pull from the same numbers you see on this page. The short model note from our index: Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video...
Related models
Models close to Qwen3.8 Flash by family, provider, and context window — each with live pricing in our catalog.