Qwen2.5 32B Instruct vs Qwen3 32B.
model vs model · prices as of
Qwen3 32B takes 1 of 5 rounds against Qwen2.5 32B Instruct. It is also the cheaper of the two to run, at $0.52 per million tokens against $1.72.
Qwen2.5 32B Instruct
$1.72/M tokens
cheapest we can explain, at Alibaba Cloud (China) · $0.224 per 1,000 pages
131k context · tool calling · open weights
Qwen3 32B
$0.52/M tokens
cheapest we can explain, at OpenRouter · $0.0648 per 1,000 pages
131k context · tool calling · open weights
List price in (lower wins)
$0.70/M
$0.70/M
List price out (lower wins)
$2.80/M
$2.80/M
Cheapest we can explain (lower wins)
$1.72/M at Alibaba Cloud (China)
$0.52/M at OpenRouter
Standard price (lower wins)
$4.90/M
$4.90/M
Cost per 1,000 pages (lower wins)
$0.224
$0.0648
Context
131k tokens
131k tokens
Open weights
yes
yes
Tool calling
yes
yes
Structured output
not stated
not stated
Reads images
no
no
Reasoning
no
yes
Released
1 Sept 2024
1 Apr 2025
Providers selling it
4
33
Price we measure against
the maker's own
the maker's own
Trains on your prompts (no is better)
not stated
not stated
Providers selling both · 2
Alibaba Cloud (China)$1.72 / $2.01
Alibaba Cloud$4.90 / $4.90
Qwen2.5 32B Instruct then Qwen3 32B, per million tokens
Go further
Add a third or fourth model, or switch the weighting, in the tool.
What moved in the last seven days
Frequently asked
- Which is cheaper, Qwen2.5 32B Instruct or Qwen3 32B?
- Qwen3 32B, at $0.52 per million tokens against $1.72 for Qwen2.5 32B Instruct, three parts input to one part output, read on 16 Sept 2026.
- Where is each one cheapest?
- Qwen2.5 32B Instruct: $1.72 at Alibaba Cloud (China). Qwen3 32B: $0.52 at OpenRouter.
- Can I buy both from one provider?
- Yes. 2 providers sell both: Alibaba Cloud (China), Alibaba Cloud. The cheapest of them for the two together is Alibaba Cloud (China).
- Which holds more context?
- Both hold 131k tokens in one call.
Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.