Requesty model pricing.
prices as of · 143 models · 154 prices
Charges what the makers charge and adds about 5% on part of the catalog, as its pricing page says. Offers EU data residency.
Requesty is a router selling 143 models, from $0.20 per million tokens for Llama 3.1 8B Instruct, read on 16 Sept 2026.
Kind of provider
Router
home country not stated
Where it runs
not stated
no regions published
How you pay
not stated
no payment terms published
Free tier
none published
nothing given away
Where Requesty's prices land, and why
Requesty publishes 154 prices we can compare, and 28 of them are below the list price for that model. 20 of those have a published reason behind them, so they rank like any other price. The other 4 sit below list on models whose weights are closed, with nothing published to account for the gap, so they are listed here in full and ranked nowhere on the site. Nobody here has read its terms yet, so treat what it says about itself as its own claim.
Requesty does not say whether it trains on what you send it. It does not say how long it keeps requests.
not yet checked: terms
Models to start with
all 154 prices ↓- DeepSeek V4 Flash 0731DeepSeek · 1M context · 97 providers sell it$0.09 / $0.1857% under listbuy at Requesty ↗
- GLM-5.2Z.ai · 1M context · 94 providers sell it$0.75 / $2.4046% under listbuy at Requesty ↗
- GLM-5.2Z.ai · 1M context · 94 providers sell it$2.10 / $6.6050% above listbuy at Requesty ↗
- DeepSeek V4 ProDeepSeek · 1M context · 84 providers sell it$1.30 / $2.60199% above listbuy at Requesty ↗
- Kimi K2.6Moonshot AI · 262k context · 82 providers sell it$0.75 / $3.5016% under listbuy at Requesty ↗
Everything Requesty sells, against the list price
Cheapest first inside each group. The share is this row against the price the lab that trained the model charges.
Cheaper, and we know why
20 pricesBelow list, with a published reason: hosting the open weights, a regional price list, credits, a stated discount.
- gpt-oss-120b
OpenAI · 131k context
Routes to a host serving the open weights; precision not disclosed.
$0.03 / $0.14in / out per M
0.84×of list price
$0.09 / $0.13in / out per M
0.89×of list price
- DeepSeek V4 Flash 0731
DeepSeek · 1M context
Routes to a host serving the open weights; precision not disclosed.
$0.09 / $0.18in / out per M
0.43×of list price
- Gemma 4 26B A4B IT
Google · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.07 / $0.34in / out per M
0.52×of list price
- Qwen2.5 72B Instruct
Alibaba Cloud · 131k context
Routes to a host serving the open weights; precision not disclosed.
$0.23 / $0.40in / out per M
0.11×of list price
$0.27 / $0.40in / out per M
0.97×of list price
- R1 Distill Llama 70B
DeepSeek · 8k context
Routes to a host serving the open weights; precision not disclosed.
$0.23 / $0.69in / out per M
0.43×of list price
- Qwen3 235B-A22B
Alibaba Cloud · 131k context
Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.
$0.20 / $0.80in / out per M
0.29×of list price
- Qwen3.5 35B-A3B
Alibaba Cloud · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.14 / $1in / out per M
0.52×of list price
- Qwen3-Next 80B-A3B (Thinking)
Alibaba Cloud · 131k context
Routes to a host serving the open weights; precision not disclosed.
$0.15 / $1.20in / out per M
0.22×of list price
- Qwen3-Coder 480B-A35B Instruct
Alibaba Cloud · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.40 / $1.60in / out per M
0.23×of list price
- Qwen3.5 27B
Alibaba Cloud · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.26 / $2.60in / out per M
1.02×of list price
- Mistral Medium (latest)
Mistral AI · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.44 / $2.20in / out per M
0.29×of list price
- Kimi K2.5
Moonshot AI · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.45 / $2.25in / out per M
0.75×of list price
$0.57 / $2.30in / out per M
0.93×of list price
$0.75 / $2.40in / out per M
0.54×of list price
- Qwen3.5 397B-A17B
Alibaba Cloud · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.49 / $3.60in / out per M
0.94×of list price
- Kimi K2.6
Moonshot AI · 262k context
Routes to a host serving the open weights; precision not disclosed.
$0.75 / $3.50in / out per M
0.84×of list price
$1.05 / $3.50in / out per M
0.77×of list price
$1.20 / $4in / out per M
0.88×of list price
Same price as the maker
130 pricesThe same price the maker charges, sold by someone else.
$0.05 / $0.05in / out per M
0.87×of list price
$0.05 / $0.20in / out per M
1.00×of list price
$0.07 / $0.14in / out per M
1.00×of list price
$0.05 / $0.20in / out per M
1.00×of list price
$0.10 / $0.30in / out per M
0.87×of list price
$0.06 / $0.44in / out per M
1.10×of list price
$0.11 / $0.33in / out per M
1.10×of list price
$0.11 / $0.33in / out per M
1.10×of list price
$0.17 / $0.17in / out per M
1.13×of list price
$0.10 / $0.40in / out per M
1.00×of list price
$0.10 / $0.40in / out per M
1.00×of list price
$0.13 / $0.38in / out per M
1.26×of list price
$0.11 / $0.44in / out per M
1.10×of list price
$0.10 / $0.50in / out per M
3.64×of list price
$0.14 / $0.40in / out per M
1.34×of list price
$0.15 / $0.50in / out per M
2.00×of list price
$0.16 / $0.47in / out per M
1.03×of list price
$0.14 / $0.58in / out per M
1.08×of list price
$0.23 / $0.40in / out per M
1.76×of list price
$0.28 / $0.28in / out per M
1.10×of list price
$0.17 / $0.66in / out per M
1.10×of list price
$0.20 / $0.60in / out per M
1.14×of list price
$0.22 / $0.66in / out per M
1.26×of list price
$0.20 / $0.80in / out per M
1.54×of list price
$0.38 / $0.40in / out per M
1.04×of list price
$0.20 / $1.10in / out per M
1.00×of list price
$0.20 / $1.15in / out per M
1.05×of list price
$0.20 / $1.25in / out per M
1.00×of list price
$0.33 / $0.99in / out per M
1.10×of list price
$0.22 / $1.32in / out per M
1.10×of list price
showing the 30 cheapest of 130 · every row in the open dataset
Cheaper, and nobody says why
4 prices · never rankedBelow list on a closed model with nothing to account for it. Shown, never ranked, never a pick.
- Qwen3 235B-A22B
Alibaba Cloud · 131k context
0.10× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.
$0.07 / $0.10in / out per M
0.06×of list price
- Qwen3 32B
Alibaba Cloud · 131k context
0.14× the list price, from one source and far below what hosting the weights costs. Listed until a second source agrees.
$0.10 / $0.30in / out per M
0.12×of list price
- Qwen3.7 Plus
Alibaba Cloud · 1M context
0.64× Alibaba Cloud's list price on a closed model, and nobody says why. Listed, never ranked.
$0.32 / $1.28in / out per M
0.50×of list price
- GPT-4o
OpenAI · 128k context
0.50× OpenAI's list price on a closed model, and nobody says why. Listed, never ranked.
$2.50 / $10in / out per M
0.58×of list price
Frequently asked
- What is Requesty?
- Charges what the makers charge and adds about 5% on part of the catalog, as its pricing page says. Offers EU data residency. It sells 143 of the models we track, and trades as Requesty.
- Is Requesty cheaper than buying from the maker?
- 28 of its 154 prices sit below the list price for the model, and 20 of those have a published reason behind them, so they can lead a ranking. 4 are listed and never ranked.
- Does Requesty train on what you send it?
- Requesty does not say whether it trains on what you send it. It does not say how long it keeps requests.
- How current are these Requesty prices?
- They come from its own public catalog, last read 16 Sept 2026. Prices exclude tax.