Llama-3.3-70B-Instruct pricing.

prices as of · 46 providers · 50 prices

Meta's Llama-3.3-70B-Instruct holds 128k tokens of context. Open weights, so anyone may host it, which is why the prices below differ so much.

Llama-3.3-70B-Instruct costs $0.10 in and $0.32 out per million tokens from OpenRouter; the cheapest price we can explain is $0.38 at NanoGPT, checked 16 Sept 2026.

tool callingopen weights128k contextreleased 6 Dec 2024

OpenRouter list price

$0.62/M tokens

$0.10 in · $0.32 out

read · 83% below the market median

Cheapest price we can explain

$0.38/M tokens

$0.05 in · $0.23 out · 39% under list

NanoGPT
View at NanoGPT

read

What it costs for your work

pages a month

OpenRouter list price

$0.0792a month

cheapest we can explain · NanoGPT

$0.0438a month

difference

$0.0354saved

A page is about 600 tokens in and 60 out, a support reply 200 in and 300 out, an email 150 in and 250 out. Nothing else is counted: no cache discount, no batch rate, no free tier.

Meta publishes the weights, so you can run Llama-3.3-70B-Instruct on a server you rent by the month. It needs about 46.5 GB of RAM on a CPU, and the cheapest plan we track with that much and a disk included is netcup VPS 8000 G12 at $46.49/mo, streaming an estimated 0.6–1.1 tok/s.

Every provider selling Llama-3.3-70B-Instruct

Four groups, decided by the price and by what the provider publishes about it. Prices are per million tokens, and the share is that row against OpenRouter's price. Meta trained Llama-3.3-70B-Instruct and sells it on no catalog we read, so the price every row is measured against is OpenRouter's.

Cheaper, and we know why

1 price

Below list, with a published reason: hosting the open weights, a regional price list, credits, a stated discount.

  • NanoGPT

    Routes to a host serving the open weights; precision not disclosed.

    131k context hereNanoGPT

    $0.05 / $0.23in / out per M

    0.61×of list price

Same price as the maker

45 prices

The same price the maker charges, sold by someone else.

  • Meganova

    The same price OpenRouter charges.

    131k context hereMeganova

    $0.10 / $0.30in / out per M

    0.97×of list price

  • OpenRouter

    The same price OpenRouter charges.

    131k context hereOpenRouter

    $0.10 / $0.32in / out per M

    1.00×of list price

  • Kilo Gateway

    The same price OpenRouter charges.

    131k context hereKilo Gateway

    $0.10 / $0.32in / out per M

    1.00×of list price

  • DeepInfra

    The same price OpenRouter charges.

    131k context herefp8 precisionDeepInfra

    $0.10 / $0.32in / out per M

    1.00×of list price

  • Nous Portal

    The same price OpenRouter charges.

    131k context hereNous Portal

    $0.10 / $0.32in / out per M

    1.00×of list price

  • Hyperbolic

    20% above OpenRouter's price.

    131k context hereHyperbolic

    $0.12 / $0.30in / out per M

    1.06×of list price

  • Helicone

    30% above OpenRouter's price.

    $0.13 / $0.39in / out per M

    1.26×of list price

  • Jiekou

    30% above OpenRouter's price.

    131k context hereJiekou

    $0.13 / $0.39in / out per M

    1.26×of list price

  • Nebius Token Factory

    30% above OpenRouter's price.

    131k context hereNebius Token Factory

    $0.13 / $0.40in / out per M

    1.27×of list price

  • Inference.net

    30% above OpenRouter's price.

    131k context hereInference.net

    $0.13 / $0.40in / out per M

    1.27×of list price

  • Nscale

    100% above OpenRouter's price.

    $0.20 / $0.20in / out per M

    1.29×of list price

  • LLM Gateway

    35% above OpenRouter's price.

    131k context hereLLM Gateway

    $0.14 / $0.40in / out per M

    1.30×of list price

  • Novita

    35% above OpenRouter's price.

    131k context hereNovita

    $0.14 / $0.40in / out per M

    1.30×of list price

  • Novita

    35% above OpenRouter's price.

    12k context herebf16 precisionNovita

    $0.14 / $0.40in / out per M

    1.30×of list price

  • TrustedRouter

    42% above OpenRouter's price.

    131k context hereTrustedRouter

    $0.14 / $0.42in / out per M

    1.37×of list price

  • DeepInfra

    130% above OpenRouter's price.

    131k context hereDeepInfra

    $0.23 / $0.40in / out per M

    1.76×of list price

  • Requesty

    130% above OpenRouter's price.

    131k context hereRequesty

    $0.23 / $0.40in / out per M

    1.76×of list price

  • AkashML

    100% above OpenRouter's price.

    131k context herefp8 precision

    $0.20 / $0.52in / out per M

    1.81×of list price

  • Merge Gateway

    120% above OpenRouter's price.

    131k context here

    $0.22 / $0.50in / out per M

    1.87×of list price

  • Parasail

    120% above OpenRouter's price.

    131k context herefp8 precisionParasail

    $0.22 / $0.50in / out per M

    1.87×of list price

  • Crusoe

    150% above OpenRouter's price.

    $0.25 / $0.75in / out per M

    2.42×of list price

  • Hugging Face Inference Providers

    490% above OpenRouter's price.

    $0.59 / $0.79in / out per M

    4.13×of list price

  • Groq

    490% above OpenRouter's price.

    131k context hereGroq

    $0.59 / $0.79in / out per M

    4.13×of list price

  • Azure Cognitive Services

    610% above OpenRouter's price.

    $0.71 / $0.71in / out per M

    4.58×of list price

  • Weights & Biases

    610% above OpenRouter's price.

    $0.71 / $0.71in / out per M

    4.58×of list price

showing the 25 cheapest of 45 · every row in the open dataset

Cheaper, and nobody says why

2 prices · never ranked

Below list on a closed model with nothing to account for it. Shown, never ranked, never a pick.

  • SambaNova

    We could not read this provider's prices reliably, so the row is listed but never ranked.

    131k context hereSambaNova

    $0.45 / $0.90in / out per M

    3.63×of list price

  • Featherless

    We could not read this provider's prices reliably, so the row is listed but never ranked.

    33k context hereFeatherless

    $0.65 / $0.75in / out per M

    4.35×of list price

Given away with limits

2 free tiers · never ranked

These providers give Llama-3.3-70B-Instruct away inside a quota of their own. A limit is not a price, so none of them counts as cheapest here or anywhere else on the site.

  • Meta
  • NVIDIA build

Frequently asked

Who sells Llama-3.3-70B-Instruct cheapest?
NanoGPT, at $0.05 in and $0.23 out per million tokens, read on 16 Sept 2026. Routes to a host serving the open weights; precision not disclosed.
Is Llama-3.3-70B-Instruct open weights?
Yes. Meta publishes the weights, so anyone may host Llama-3.3-70B-Instruct, which is why 46 providers sell it and why their prices differ so much.
How much context does Llama-3.3-70B-Instruct hold?
128,000 tokens in one call, which is about 128k, and it can call your own functions.
How does Llama-3.3-70B-Instruct compare with the rest of the market?
At $0.62 per million tokens, OpenRouter's price is 83% below the $3.55 median across the 337 text models we track.
More from MetaLlama 3.1 8B InstructLlama 3.2 3B InstructMuse Spark 1.2Muse Glimmer 30BSimilar contextgpt-oss-120bDeepSeek V3.2gpt-oss-20bR1 0528