Requesty model pricing.

prices as of · 143 models · 154 prices

Charges what the makers charge and adds about 5% on part of the catalog, as its pricing page says. Offers EU data residency.

Requesty is a router selling 143 models, from $0.20 per million tokens for Llama 3.1 8B Instruct, read on 16 Sept 2026.

Requesty prices

Kind of provider

Router

home country not stated

Where it runs

not stated

no regions published

How you pay

not stated

no payment terms published

Free tier

none published

nothing given away

Where Requesty's prices land, and why

Requesty publishes 154 prices we can compare, and 28 of them are below the list price for that model. 20 of those have a published reason behind them, so they rank like any other price. The other 4 sit below list on models whose weights are closed, with nothing published to account for the gap, so they are listed here in full and ranked nowhere on the site. Nobody here has read its terms yet, so treat what it says about itself as its own claim.

Requesty does not say whether it trains on what you send it. It does not say how long it keeps requests.

not yet checked: terms

Models to start with

all 154 prices ↓

Everything Requesty sells, against the list price

Cheapest first inside each group. The share is this row against the price the lab that trained the model charges.

Cheaper, and we know why

20 prices

Below list, with a published reason: hosting the open weights, a regional price list, credits, a stated discount.

  • gpt-oss-120b

    OpenAI · 131k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.03 / $0.14in / out per M

    0.84×of list price

  • Qwen3.5-9B

    Alibaba Cloud · 262k context

    10% under the list price, an ordinary reseller margin.

    $0.09 / $0.13in / out per M

    0.89×of list price

  • DeepSeek V4 Flash 0731

    DeepSeek · 1M context

    Routes to a host serving the open weights; precision not disclosed.

    $0.09 / $0.18in / out per M

    0.43×of list price

  • Gemma 4 26B A4B IT

    Google · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.07 / $0.34in / out per M

    0.52×of list price

  • Qwen2.5 72B Instruct

    Alibaba Cloud · 131k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.23 / $0.40in / out per M

    0.11×of list price

  • DeepSeek V3.2

    DeepSeek · 164k context

    4% under the list price, an ordinary reseller margin.

    $0.27 / $0.40in / out per M

    0.97×of list price

  • R1 Distill Llama 70B

    DeepSeek · 8k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.23 / $0.69in / out per M

    0.43×of list price

  • Qwen3 235B-A22B

    Alibaba Cloud · 131k context

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    $0.20 / $0.80in / out per M

    0.29×of list price

  • Qwen3.5 35B-A3B

    Alibaba Cloud · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.14 / $1in / out per M

    0.52×of list price

  • Qwen3-Next 80B-A3B (Thinking)

    Alibaba Cloud · 131k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.15 / $1.20in / out per M

    0.22×of list price

  • Qwen3-Coder 480B-A35B Instruct

    Alibaba Cloud · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.40 / $1.60in / out per M

    0.23×of list price

  • Qwen3.5 27B

    Alibaba Cloud · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.26 / $2.60in / out per M

    1.02×of list price

  • Mistral Medium (latest)

    Mistral AI · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.44 / $2.20in / out per M

    0.29×of list price

  • Kimi K2.5

    Moonshot AI · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.45 / $2.25in / out per M

    0.75×of list price

  • Kimi K2 0905

    Moonshot AI · 262k context

    5% under the list price, an ordinary reseller margin.

    $0.57 / $2.30in / out per M

    0.93×of list price

  • GLM-5.2

    Z.ai · 1M context

    Routes to a host serving the open weights; precision not disclosed.

    $0.75 / $2.40in / out per M

    0.54×of list price

  • Qwen3.5 397B-A17B

    Alibaba Cloud · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.49 / $3.60in / out per M

    0.94×of list price

  • Kimi K2.6

    Moonshot AI · 262k context

    Routes to a host serving the open weights; precision not disclosed.

    $0.75 / $3.50in / out per M

    0.84×of list price

  • GLM-5.1

    Z.ai · 200k context

    Routes to a host serving the open weights; precision not disclosed.

    $1.05 / $3.50in / out per M

    0.77×of list price

  • GLM-5.3

    Z.ai · 1M context

    Routes to a host serving the open weights; precision not disclosed.

    $1.20 / $4in / out per M

    0.88×of list price

Same price as the maker

130 prices

The same price the maker charges, sold by someone else.

showing the 30 cheapest of 130 · every row in the open dataset

Cheaper, and nobody says why

4 prices · never ranked

Below list on a closed model with nothing to account for it. Shown, never ranked, never a pick.

  • Qwen3 235B-A22B

    Alibaba Cloud · 131k context

    0.10× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    $0.07 / $0.10in / out per M

    0.06×of list price

  • Qwen3 32B

    Alibaba Cloud · 131k context

    0.14× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.

    $0.10 / $0.30in / out per M

    0.12×of list price

  • Qwen3.7 Plus

    Alibaba Cloud · 1M context

    0.64× Alibaba Cloud's list price on a closed model, and nobody says why. Listed, never ranked.

    $0.32 / $1.28in / out per M

    0.50×of list price

  • GPT-4o

    OpenAI · 128k context

    0.50× OpenAI's list price on a closed model, and nobody says why. Listed, never ranked.

    $2.50 / $10in / out per M

    0.58×of list price

Frequently asked

What is Requesty?
Charges what the makers charge and adds about 5% on part of the catalog, as its pricing page says. Offers EU data residency. It sells 143 of the models we track, and trades as Requesty.
Is Requesty cheaper than buying from the maker?
28 of its 154 prices sit below the list price for the model, and 20 of those have a published reason behind them, so they can lead a ranking. 4 are listed and never ranked.
Does Requesty train on what you send it?
Requesty does not say whether it trains on what you send it. It does not say how long it keeps requests.
How current are these Requesty prices?
They come from its own public catalog, last read 16 Sept 2026. Prices exclude tax.