Qwen2.5 14B Instruct vs WizardLM-2 8x22B.

model vs model · prices as of

Qwen2.5 14B Instruct takes 4 of 5 rounds against WizardLM-2 8x22B. It is also the cheaper of the two to run, at $0.86 per million tokens against $1.92.

Alibaba Cloudwins 4 rounds

Qwen2.5 14B Instruct

$0.86/M tokens

cheapest we can explain, at Alibaba Cloud (China) · $0.112 per 1,000 pages

131k context · tool calling · open weights

Microsoft

WizardLM-2 8x22B

$1.92/M tokens

cheapest we can explain, at DeepInfra · $0.317 per 1,000 pages

66k context · open weights

DeepInfra

List price in (lower wins)

$0.35/M

$0.62/M

List price out (lower wins)

$1.40/M

$0.62/M

Cheapest we can explain (lower wins)

$0.86/M at Alibaba Cloud (China)

$1.92/M at DeepInfra

Standard price (lower wins)

$2.45/M

$2.48/M

Cost per 1,000 pages (lower wins)

$0.112

$0.317

Context

131k tokens

66k tokens

Open weights

yes

yes

Tool calling

yes

no

Structured output

not stated

yes

Reads images

no

no

Reasoning

no

no

Released

1 Sept 2024

16 Apr 2024

Providers selling it

4

8

Price we measure against

the maker's own

OpenRouter

Trains on your prompts (no is better)

not stated

not stated

Providers selling both · 0

No provider we read sells both, so running the two means two accounts.

Go further

Add a third or fourth model, or switch the weighting, in the tool.

Frequently asked

Which is cheaper, Qwen2.5 14B Instruct or WizardLM-2 8x22B?
Qwen2.5 14B Instruct, at $0.86 per million tokens against $1.92 for WizardLM-2 8x22B, three parts input to one part output, read on 16 Sept 2026.
Where is each one cheapest?
Qwen2.5 14B Instruct: $0.86 at Alibaba Cloud (China). WizardLM-2 8x22B: $1.92 at DeepInfra.
Can I buy both from one provider?
No provider we read sells both, so running the two means two accounts and two bills.
Which holds more context?
Qwen2.5 14B Instruct, at 131k tokens against 66k.

Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.

Related head-to-headsERNIE 4.5 VL 424B A47B vs WizardLM-2 8x22BQwen2.5 14B Instruct vs Qwen3 14BQwen3 14B vs WizardLM-2 8x22BMorecompare any two modelsevery model we track