Llama-3.3-70B-Instruct pricing.
prices as of · 46 providers · 50 prices
Meta's Llama-3.3-70B-Instruct holds 128k tokens of context. Open weights, so anyone may host it, which is why the prices below differ so much.
Llama-3.3-70B-Instruct costs $0.10 in and $0.32 out per million tokens from OpenRouter; the cheapest price we can explain is $0.38 at NanoGPT, checked 16 Sept 2026.
OpenRouter list price
$0.62/M tokens
$0.10 in · $0.32 out
read · 83% below the market median
Cheapest price we can explain
$0.38/M tokens
$0.05 in · $0.23 out · 39% under list
read
What it costs for your work
OpenRouter list price
$0.0792a month
cheapest we can explain · NanoGPT
$0.0438a month
difference
$0.0354saved
A page is about 600 tokens in and 60 out, a support reply 200 in and 300 out, an email 150 in and 250 out. Nothing else is counted: no cache discount, no batch rate, no free tier.
Run it yourself
every plan that fits Llama-3.3-70B-Instruct →Meta publishes the weights, so you can run Llama-3.3-70B-Instruct on a server you rent by the month. It needs about 46.5 GB of RAM on a CPU, and the cheapest plan we track with that much and a disk included is netcup VPS 8000 G12 at $46.49/mo, streaming an estimated 0.6–1.1 tok/s.
Every provider selling Llama-3.3-70B-Instruct
Four groups, decided by the price and by what the provider publishes about it. Prices are per million tokens, and the share is that row against OpenRouter's price. Meta trained Llama-3.3-70B-Instruct and sells it on no catalog we read, so the price every row is measured against is OpenRouter's.
Cheaper, and we know why
1 priceBelow list, with a published reason: hosting the open weights, a regional price list, credits, a stated discount.
Same price as the maker
45 pricesThe same price the maker charges, sold by someone else.
$0.10 / $0.30in / out per M
0.97×of list price
$0.10 / $0.32in / out per M
1.00×of list price
$0.10 / $0.32in / out per M
1.00×of list price
$0.10 / $0.32in / out per M
1.00×of list price
$0.10 / $0.32in / out per M
1.00×of list price
$0.12 / $0.30in / out per M
1.06×of list price
Helicone30% above OpenRouter's price.
$0.13 / $0.39in / out per M
1.26×of list price
$0.13 / $0.39in / out per M
1.26×of list price
$0.13 / $0.40in / out per M
1.27×of list price
$0.13 / $0.40in / out per M
1.27×of list price
$0.20 / $0.20in / out per M
1.29×of list price
$0.14 / $0.40in / out per M
1.30×of list price
$0.14 / $0.40in / out per M
1.30×of list price
$0.14 / $0.40in / out per M
1.30×of list price
$0.14 / $0.42in / out per M
1.37×of list price
$0.23 / $0.40in / out per M
1.76×of list price
$0.23 / $0.40in / out per M
1.76×of list price
AkashML100% above OpenRouter's price.
131k context herefp8 precision$0.20 / $0.52in / out per M
1.81×of list price
$0.22 / $0.50in / out per M
1.87×of list price
$0.22 / $0.50in / out per M
1.87×of list price
$0.25 / $0.75in / out per M
2.42×of list price
Hugging Face Inference Providers490% above OpenRouter's price.
131k context hereHugging Face Inference Providers ↗$0.59 / $0.79in / out per M
4.13×of list price
$0.59 / $0.79in / out per M
4.13×of list price
Azure Cognitive Services610% above OpenRouter's price.
$0.71 / $0.71in / out per M
4.58×of list price
Weights & Biases610% above OpenRouter's price.
$0.71 / $0.71in / out per M
4.58×of list price
showing the 25 cheapest of 45 · every row in the open dataset
Cheaper, and nobody says why
2 prices · never rankedBelow list on a closed model with nothing to account for it. Shown, never ranked, never a pick.
SambaNovaWe could not read this provider's prices reliably, so the row is listed but never ranked.
131k context hereSambaNova ↗$0.45 / $0.90in / out per M
3.63×of list price
FeatherlessWe could not read this provider's prices reliably, so the row is listed but never ranked.
33k context hereFeatherless ↗$0.65 / $0.75in / out per M
4.35×of list price
Given away with limits
2 free tiers · never rankedThese providers give Llama-3.3-70B-Instruct away inside a quota of their own. A limit is not a price, so none of them counts as cheapest here or anywhere else on the site.
Meta
NVIDIA build
Frequently asked
- Who sells Llama-3.3-70B-Instruct cheapest?
- NanoGPT, at $0.05 in and $0.23 out per million tokens, read on 16 Sept 2026. Routes to a host serving the open weights; precision not disclosed.
- Is Llama-3.3-70B-Instruct open weights?
- Yes. Meta publishes the weights, so anyone may host Llama-3.3-70B-Instruct, which is why 46 providers sell it and why their prices differ so much.
- How much context does Llama-3.3-70B-Instruct hold?
- 128,000 tokens in one call, which is about 128k, and it can call your own functions.
- How does Llama-3.3-70B-Instruct compare with the rest of the market?
- At $0.62 per million tokens, OpenRouter's price is 83% below the $3.55 median across the 337 text models we track.