Qwen3 14B
The cheapest provider for Qwen3 14B is NextBit at $0.13 per 1M tokens (blended), charging $0.10 per 1M input and $0.22 per 1M output tokens. 3 providers serve it; the most expensive charges 3.1× as much. Context window: 128K tokens. Prices tracked since Oct 3, 2026.
Best price
$0.13
at NextBit
Input / output
$0.10 / $0.22
per 1M tokens
Providers
3
3.1× spread
7 days
—
best price
30 days
—
best price
Context
128K
tokens
Providers
Active offers ranked by blended price. Bars are relative to the most expensive.
| Provider | Input | Output | Blended | vs cheapest | Context | Uptime 1d | Price since |
|---|---|---|---|---|---|---|---|
NextBitCheapest int4 | $0.10 | $0.22 | $0.13 | best | |||
| fp8 | $0.12 | $0.24 | $0.15 | +15% | |||
| $0.228 | $0.91 | $0.398 | +206% |
- NextBitCheapest$0.13$0.10 in · $0.22 outbest40K ctx99.3% uptimeint4
- $0.15$0.12 in · $0.24 out+15%40K ctx97.0% uptimefp8
- $0.398$0.228 in · $0.91 out+206%128K ctx99.2% uptime
Blended = (3 × input + output) / 4.
Price history
Hover to compare providers at any point in time. Click a legend item to hide it.
USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.
Change log
Every recorded change for this model since Oct 3, 2026.
No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.
About this model
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
- Released
- Apr 28, 2025
- First tracked
- Oct 3, 2026
- Canonical slug
- qwen/qwen3-14b-04-28
- Weights
- Qwen/Qwen3-14B