Qwen3 32B pricing.

prices as of · 33 providers · 37 prices

Alibaba Cloud's Qwen3 32B holds 131k tokens of context and has a reasoning mode. Open weights, so anyone may host it, which is why the prices below differ so much.

Qwen3 32B costs $0.70 in and $2.80 out per million tokens from Alibaba Cloud; the cheapest price we can explain is $0.52 at OpenRouter, checked 16 Sept 2026.

tool callingreasoningopen weights131k contextreleased 1 Apr 2025

Alibaba Cloud list price

$4.90/M tokens

$0.70 in · $2.80 out

read · 38% above the market median

Cheapest price we can explain

$0.52/M tokens

$0.08 in · $0.28 out · 89% under list

OpenRouter
View at OpenRouter

read

What it costs for your work

pages a month

Alibaba Cloud list price

$0.588a month

cheapest we can explain · OpenRouter

$0.0648a month

difference

$0.523saved

A page is about 600 tokens in and 60 out, a support reply 200 in and 300 out, an email 150 in and 250 out. Nothing else is counted: no cache discount, no batch rate, no free tier.

Alibaba Cloud publishes the weights, so you can run Qwen3 32B on a server you rent by the month. It needs about 23.3 GB of RAM on a CPU, and the cheapest plan we track with that much and a disk included is netcup VPS 4000 G12 at $31.43/mo, streaming an estimated 1.2–2.4 tok/s.

Every provider selling Qwen3 32B

Four groups, decided by the price and by what the provider publishes about it. Prices are per million tokens, and the share is that row against Alibaba Cloud's price.

The maker's list price

1 price

What the lab that trained the model charges on its own API.

  • Alibaba Cloud

    The price Alibaba Cloud charges for the model it trained. Every row below is measured against it.

    $0.70 / $2.80in / out per M

    1.00×of list price

Cheaper, and we know why

21 prices

Below list, with a published reason: hosting the open weights, a regional price list, credits, a stated discount.

  • OpenRouter

    Routes to a host serving the open weights; precision not disclosed.

    $0.08 / $0.28in / out per M

    0.11×of list price

  • DeepInfra

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    41k context hereDeepInfra

    $0.08 / $0.28in / out per M

    0.11×of list price

  • Kilo Gateway

    Routes to a host serving the open weights; precision not disclosed.

    41k context hereKilo Gateway

    $0.08 / $0.28in / out per M

    0.11×of list price

  • DeepInfra

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    41k context herefp8 precisionDeepInfra

    $0.08 / $0.28in / out per M

    0.11×of list price

  • Nous Portal

    Resells OpenRouter's catalog at 0.80× the listed price, funded by a subscription whose credits carry a 10% bonus.

    $0.08 / $0.28in / out per M

    0.11×of list price

  • OVHcloud AI Endpoints

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $0.09 / $0.25in / out per M

    0.11×of list price

  • Eden AI

    Alibaba's China-region price list, which is lower than the international one. A region, not a deal.

    33k context hereEden AI

    $0.09 / $0.25in / out per M

    0.11×of list price

  • Cortecs

    Routes to a host serving the open weights; precision not disclosed.

    32k context hereCortecs

    $0.09 / $0.32in / out per M

    0.12×of list price

  • Novita

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    41k context herefp8 precisionNovita

    $0.10 / $0.45in / out per M

    0.15×of list price

  • Jiekou

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    41k context herefp8 precisionJiekou

    $0.10 / $0.45in / out per M

    0.15×of list price

  • NEAR AI Cloud

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    128k context hereNEAR AI Cloud

    $0.11 / $0.46in / out per M

    0.16×of list price

  • Phala (RedPill)

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    41k context herePhala (RedPill)

    $0.12 / $0.50in / out per M

    0.18×of list price

  • SiliconFlow (China)

    SiliconFlow's China-region catalog, priced in the domestic market.

    131k context hereSiliconFlow (China)

    $0.14 / $0.57in / out per M

    0.20×of list price

  • SiliconFlow

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    131k context hereSiliconFlow

    $0.14 / $0.57in / out per M

    0.20×of list price

  • SiliconFlow

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    fp8 precisionSiliconFlow

    $0.14 / $0.57in / out per M

    0.20×of list price

  • Merge Gateway

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $0.15 / $0.60in / out per M

    0.21×of list price

  • Helicone

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $0.29 / $0.59in / out per M

    0.30×of list price

  • Hugging Face Inference Providers

    Routes to a host serving the open weights; precision not disclosed.

    $0.29 / $0.59in / out per M

    0.30×of list price

  • Groq

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    131k context hereGroq

    $0.29 / $0.59in / out per M

    0.30×of list price

  • LLM Gateway

    Routes to a host serving the open weights; precision not disclosed.

    33k context hereLLM Gateway

    $0.36 / $0.87in / out per M

    0.40×of list price

  • Alibaba Cloud (China)

    Alibaba's China-region Model Studio, which prices the Qwen models below the international list. A region, not a deal.

    $0.29 / $1.15in / out per M

    0.41×of list price

Same price as the maker

1 price

The same price the maker charges, sold by someone else.

Cheaper, and nobody says why

13 prices · never ranked

Below list on a closed model with nothing to account for it. Shown, never ranked, never a pick.

  • Lambda

    0.07× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    fp8 precisionLambda

    $0.05 / $0.10in / out per M

    0.05×of list price

  • Nscale

    0.11× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereNscale

    $0.08 / $0.25in / out per M

    0.10×of list price

  • TrustedRouter

    0.12× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereTrustedRouter

    $0.08 / $0.26in / out per M

    0.11×of list price

  • FastRouter

    0.11× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereFastRouter

    $0.08 / $0.28in / out per M

    0.11×of list price

  • Abacus

    0.13× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    $0.09 / $0.29in / out per M

    0.11×of list price

  • NanoGPT

    0.14× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereNanoGPT

    $0.10 / $0.30in / out per M

    0.12×of list price

  • Nebius Token Factory

    0.14× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereNebius Token Factory

    $0.10 / $0.30in / out per M

    0.12×of list price

  • Requesty

    0.14× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereRequesty

    $0.10 / $0.30in / out per M

    0.12×of list price

  • Chutes

    0.15× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context hereChutes

    $0.10 / $0.42in / out per M

    0.15×of list price

  • Chutes

    0.15× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    41k context herefp8 precisionChutes

    $0.10 / $0.42in / out per M

    0.15×of list price

  • Featherless

    We could not read this provider's prices reliably, so the row is listed but never ranked.

    33k context hereFeatherless

    $0.10 / $0.49in / out per M

    0.16×of list price

  • Pioneer

    A router whose catalog we cannot yet read consistently.

    $0.16 / $0.64in / out per M

    0.23×of list price

  • SambaNova

    We could not read this provider's prices reliably, so the row is listed but never ranked.

    8k context hereSambaNova

    $0.40 / $0.80in / out per M

    0.41×of list price

Given away with limits

1 free tiers · never ranked

These providers give Qwen3 32B away inside a quota of their own. A limit is not a price, so none of them counts as cheapest here or anywhere else on the site.

  • iFlow

Frequently asked

Who sells Qwen3 32B cheapest?
OpenRouter, at $0.08 in and $0.28 out per million tokens, read on 16 Sept 2026. Routes to a host serving the open weights; precision not disclosed.
Is Qwen3 32B open weights?
Yes. Alibaba Cloud publishes the weights, so anyone may host Qwen3 32B, which is why 33 providers sell it and why their prices differ so much.
How much context does Qwen3 32B hold?
131,072 tokens in one call, which is about 131k, and it can call your own functions.
How does Qwen3 32B compare with the rest of the market?
At $4.90 per million tokens, Alibaba Cloud's price is 38% above the $3.55 median across the 337 text models we track.
More from Alibaba CloudQwen3.8 27BQwen3.5 397B-A17BQwen3.6 35B-A3BQwen3 235B-A22BSimilar contextgpt-oss-120bDeepSeek V3.2Llama-3.3-70B-Instructgpt-oss-20bRanked in#4 in Biggest savings against the list price