DeepSeek V4.1 Flash vs Qwen2.5-VL 7B Instruct.
model vs model · prices as of
DeepSeek V4.1 Flash takes 3 of 5 rounds against Qwen2.5-VL 7B Instruct. It is also the cheaper of the two to run, at $0.70 per million tokens against $1.58.
DeepSeek V4.1 Flash
$0.70/M tokens
cheapest we can explain, at AMD · $0.101 per 1,000 pages
1M context · tool calling · reads images · open weights
Qwen2.5-VL 7B Instruct
$1.58/M tokens
cheapest we can explain, at Alibaba Cloud (China) · $0.215 per 1,000 pages
131k context · tool calling · reads images · open weights
List price in (lower wins)
$0.30/M
$0.35/M
List price out (lower wins)
$1.20/M
$1.05/M
Cheapest we can explain (lower wins)
$0.70/M at AMD
$1.58/M at Alibaba Cloud (China)
Standard price (lower wins)
$2.10/M
$2.10/M
Cost per 1,000 pages (lower wins)
$0.101
$0.215
Context
1M tokens
131k tokens
Open weights
yes
yes
Tool calling
yes
yes
Structured output
yes
not stated
Reads images
yes
yes
Reasoning
yes
no
Released
10 Sept 2026
1 Sept 2024
Providers selling it
58
3
Price we measure against
the maker's own
the maker's own
Trains on your prompts (no is better)
not stated
not stated
Providers selling both · 1
Alibaba Cloud$2.10 / $2.10
DeepSeek V4.1 Flash then Qwen2.5-VL 7B Instruct, per million tokens
Go further
Add a third or fourth model, or switch the weighting, in the tool.
What moved in the last seven days
- DeepSeek V4.1 Flash
DeepSeek$0.15 / $0.60 $0.30 / $1.20100% dearer · - DeepSeek V4.1 Flash
io.net Intelligence$0.31 / $1.24 $0.31 / $1.221.0% cheaper · - DeepSeek V4.1 Flash
DeepSeekDeepSeek's list price, in and out per million tokens$0.15 / $0.60 $0.30 / $1.20100% dearer · - DeepSeek V4.1 Flash
Morph$0.24 / $0.96 $0.15 / $0.5938% cheaper · - DeepSeek V4.1 Flash
Nous Portal$0.15 / $0.60 $0.15 / $0.591.0% cheaper · - DeepSeek V4.1 Flash
Novita$0.30 / $1.20 $0.28 / $1.145% cheaper · - DeepSeek V4.1 Flash
OpenRouter$0.15 / $0.60 $0.30 / $1.20100% dearer · - DeepSeek V4.1 Flash
Requesty$0.15 / $0.60 $0.50 / $1.50186% dearer · - DeepSeek V4.1 Flash
Wafer$0.30 / $1.20 $0.20 / $0.6043% cheaper · - DeepSeek V4.1 Flash
Charm Hyper$0.33 / $1.31 $0.30 / $1.208% cheaper · - DeepSeek V4.1 Flash
Eden AI$0.50 / $1.50 $0.30 / $1.2030% cheaper · - DeepSeek V4.1 Flash
Vercel AI Gateway$0.15 / $0.60 $0.30 / $1.20100% dearer · - DeepSeek V4.1 Flash
GMI Cloud$0.30 / $1.20 $0.28 / $1.145% cheaper · - DeepSeek V4.1 Flash
Kenari$0.35 / $0.70 $0.01 / $0.0396% cheaper · - DeepSeek V4.1 Flash
TrustedRouter$0.16 / $0.63 $0.04 / $0.0881% cheaper · - DeepSeek V4.1 Flash
DeepInfra$0.30 / $1.20 $0.20 / $0.6043% cheaper ·
Frequently asked
- Which is cheaper, DeepSeek V4.1 Flash or Qwen2.5-VL 7B Instruct?
- DeepSeek V4.1 Flash, at $0.70 per million tokens against $1.58 for Qwen2.5-VL 7B Instruct, three parts input to one part output, read on 18 Sept 2026.
- Where is each one cheapest?
- DeepSeek V4.1 Flash: $0.70 at AMD. Qwen2.5-VL 7B Instruct: $1.58 at Alibaba Cloud (China).
- Can I buy both from one provider?
- Yes. One provider sells both: Alibaba Cloud. The cheapest of them for the two together is Alibaba Cloud.
- Which holds more context?
- DeepSeek V4.1 Flash, at 1M tokens against 131k.
Prices exclude tax and are per million tokens, read 18 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.