AlibabaQwenTool use

Qwen3 14B pricing

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for tasks like math, programming, and logical inference, and a "non-thinking" mode for general-purpose conversation. The model is fine-tuned for instruction-following, agent tool use, creative writing, and multilingual tasks across 100+ languages and dialects. It natively handles 32K token contexts a. Below you will find 5 current rows with input and output dollars per million. Right now the lowest input is $0.050 and the lowest output is $0.200.

41K context·5 providers·verified Sep 5, 2026
Best input$0.050per 1M tokens · Wandb
Best output$0.200per 1M tokens · Fireworks AI

Pricing across providers

All figures are list prices per million tokens unless a column says otherwise. 5 offers are listed for Qwen3 14B. Best input in this view: Wandb.

D
DeepInfra
Input / 1M
$0.120
Output / 1M
$0.240
FA
Fireworks AI
Input / 1M
$0.200
Output / 1M
$0.200
N
Nebius
Input / 1M
$0.080
Output / 1M
$0.240
O
Openrouter
Input / 1M
$0.120
Output / 1M
$0.240
W
Wandb
Input / 1M
$0.050
Output / 1M
$0.220

Input vs output · per provider

Cost calculator

Use this block to stress test Qwen3 14B cost without a spreadsheet. All estimates come from public list rates in this page.

Provider

In: $0.120/M·Out: $0.240/M

0.012000¢ / req

0.012000¢ / req

Daily
$2.40
Monthly
$72
Annual
$876

Model specifications

These fields describe Qwen3 14B as we store it (Family: Qwen. source: Alibaba). They sit next to price so buyers can check limits and tools in one place.

Context window
40,960 tokens
Max output
40,960 tokens
Vision (images)
No
Tool / function calling
Yes
Streaming
No
Released
Apr 2025
Primary provider
Alibaba
Model family
Qwen

Compare Qwen3 14B

Open a pair page to see Qwen3 14B next to another model with a shared provider matrix. 6 shortcuts below.

Locked

Compare with

Pick a model on both sides.

Popular Qwen3 14B comparisons

Frequently asked questions

Answers pull from the same numbers you see on this page. The short model note from our index: Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for tasks like math, ...

Yes. Qwen3 14B is available on DeepInfra, Fireworks AI, Nebius, Openrouter, Wandb.

Related models

Models close to Qwen3 14B by family, provider, and context window — each with live pricing in our catalog.