Qwen 2.5 14B API Pricing
Open-weight · 128K context · Available from 7 hosts. Cheapest: Novita at $0.08 input / $0.08 output per 1M tokens.
Among Alibaba's 10 open-weight models tracked here, Qwen 2.5 14B ($0.1/$0.1 per 1M tokens) is above Qwen 2.5 7B ($0.05/$0.05) and below Qwen 2.5 32B ($0.2/$0.2).
Qwen 2.5 14B price by host (7 providers)
| Host | Input /1M | Output /1M | In+Out |
|---|---|---|---|
| Novita | $0.08 | $0.08 | $0.16 |
| DeepInfra | $0.085 | $0.085 | $0.17 |
| Groq | $0.095 | $0.095 | $0.19 |
| Alibaba | $0.1 | $0.1 | $0.2 |
| Together | $0.1 | $0.1 | $0.2 |
| OpenRouter | $0.1 | $0.1 | $0.2 |
| Fireworks | $0.105 | $0.105 | $0.21 |
Prices per 1M tokens, USD. Snapshot 2026-07 — verify with the provider before relying on them.
Qwen 2.5 14B cost calculator
Frequently asked
Qwen 2.5 14B costs $0.08 per 1M input tokens and $0.08 per 1M output tokens at its cheapest host (Novita).
Qwen 2.5 14B supports a context window of about 128K tokens. Remember every token in the context is billed on each call.
Across 7 hosts, Novita is currently cheapest at $0.08/$0.08 per 1M tokens, versus Fireworks at $0.105/$0.105. Prices change often — use the comparison table above.