HomeAI Models › Qwen 2.5 32B

Qwen 2.5 32B API Pricing

Open-weight · 128K context · Available from 7 hosts. Cheapest: Novita at $0.16 input / $0.16 output per 1M tokens.

Qwen 2.5 32B price by host (7 providers)

HostInput /1MOutput /1MIn+Out
Novita$0.16$0.16$0.32
DeepInfra$0.17$0.17$0.34
Groq$0.19$0.19$0.38
Alibaba$0.2$0.2$0.4
Together$0.2$0.2$0.4
OpenRouter$0.2$0.2$0.4
Fireworks$0.21$0.21$0.42

Prices per 1M tokens, USD. Snapshot 2026-07 — verify with the provider before relying on them.

Qwen 2.5 32B cost calculator

Frequently asked

How much does Qwen 2.5 32B cost per 1M tokens?

Qwen 2.5 32B costs $0.16 per 1M input tokens and $0.16 per 1M output tokens at its cheapest host (Novita).

What is Qwen 2.5 32B's context window?

Qwen 2.5 32B supports a context window of about 128K tokens. Remember every token in the context is billed on each call.

Where is Qwen 2.5 32B cheapest to run?

Across 7 hosts, Novita is currently cheapest at $0.16/$0.16 per 1M tokens, versus Fireworks at $0.21/$0.21. Prices change often — use the comparison table above.

Related models

Compare all 387 models →