Inkling pricing.

prices as of · 23 providers · 27 prices

Thinking Machines' Inkling holds 1M tokens of context, reads images as well as text and has a reasoning mode. Open weights, so anyone may host it, which is why the prices below differ so much.

Inkling costs $1.87 in and $4.68 out per million tokens from Thinking Machines; the cheapest price we can explain is $6.90 at DeepInfra, checked 16 Sept 2026.

tool callingreads imagesreasoningopen weights1M contextreleased 17 Jul 2026

Thinking Machines list price

$10.29/M tokens

$1.87 in · $4.68 out

read · 190% above the market median

Cheapest price we can explain

$6.90/M tokens

$0.95 in · $4.05 out · 33% under list

DeepInfra
View at DeepInfra

read

What it costs for your work

pages a month

Thinking Machines list price

$1.40a month

cheapest we can explain · DeepInfra

$0.813a month

difference

$0.59saved

A page is about 600 tokens in and 60 out, a support reply 200 in and 300 out, an email 150 in and 250 out. Nothing else is counted: no cache discount, no batch rate, no free tier.

Every provider selling Inkling

Four groups, decided by the price and by what the provider publishes about it. Prices are per million tokens, and the share is that row against Thinking Machines's price.

Cheaper, and we know why

21 prices

Below list, with a published reason: hosting the open weights, a regional price list, credits, a stated discount.

  • DeepInfra

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    524k context hereDeepInfra

    $0.95 / $4.05in / out per M

    0.67×of list price

  • Kilo Gateway

    Routes to a host serving the open weights; precision not disclosed.

    524k context hereKilo Gateway

    $0.95 / $4.05in / out per M

    0.67×of list price

  • DeepInfra

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    524k context herefp8 precisionDeepInfra

    $0.95 / $4.05in / out per M

    0.67×of list price

  • Nous Portal

    Resells OpenRouter's catalog at 0.80× the listed price, funded by a subscription whose credits carry a 10% bonus.

    $0.95 / $4.05in / out per M

    0.67×of list price

  • OpenRouter

    Routes to a host serving the open weights; precision not disclosed.

    $1 / $4.05in / out per M

    0.69×of list price

  • Hugging Face Inference Providers

    Routes to a host serving the open weights; precision not disclosed.

    $1 / $4.05in / out per M

    0.69×of list price

  • Baseten

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $1 / $4.05in / out per M

    0.69×of list price

  • Vercel AI Gateway

    Routes to a host serving the open weights; precision not disclosed.

    256k context hereVercel AI Gateway

    $1 / $4.05in / out per M

    0.69×of list price

  • Fireworks AI

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $1 / $4.05in / out per M

    0.69×of list price

  • Merge Gateway

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    1M context here

    $1 / $4.05in / out per M

    0.69×of list price

  • Neon

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $1 / $4.05in / out per M

    0.69×of list price

  • Together AI

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    524k context hereTogether AI

    $1 / $4.05in / out per M

    0.69×of list price

  • Eden AI

    Routes to a host serving the open weights; precision not disclosed.

    524k context hereEden AI

    $1 / $4.05in / out per M

    0.69×of list price

  • Baseten

    Serves the open weights at fp8, a lower precision than the maker's, so not quite the same product.

    fp8 precisionBaseten

    $1 / $4.05in / out per M

    0.69×of list price

  • Poe

    A consumer subscription with an API bolted on. Points convert to dollars, so the rate moves when the conversion does.

    256k context herePoe

    $1.01 / $4.09in / out per M

    0.69×of list price

  • TrustedRouter

    Routes to a host serving the open weights; precision not disclosed.

    524k context hereTrustedRouter

    $1.00 / $4.27in / out per M

    0.71×of list price

  • Charm Hyper

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    $1.09 / $4.41in / out per M

    0.75×of list price

  • Modal

    Serves the open weights at nvfp4, a lower precision than the maker's, so not quite the same product.

    nvfp4 precision

    $1.20 / $5in / out per M

    0.84×of list price

  • Venice

    Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.

    524k context hereVenice

    $1.25 / $5.06in / out per M

    0.86×of list price

  • NanoGPT

    9% under the list price, an ordinary reseller margin.

    1M context hereNanoGPT

    $1.70 / $6.89in / out per M

    1.17×of list price

  • NanoGPT

    9% under the list price, an ordinary reseller margin.

    1M context hereNanoGPT

    $1.70 / $6.89in / out per M

    1.17×of list price

Same price as the maker

5 prices

The same price the maker charges, sold by someone else.

  • Impossibl

    The same price Thinking Machines charges.

    66k context hereImpossibl

    $1.87 / $4.68in / out per M

    1.00×of list price

  • Requesty

    The same price Thinking Machines charges.

    66k context hereRequesty

    $1.87 / $4.68in / out per M

    1.00×of list price

  • LLMTR

    The same price Thinking Machines charges.

    262k context hereLLMTR

    $1.87 / $4.68in / out per M

    1.00×of list price

  • Thinking Machines

    The same price Thinking Machines charges.

    66k context here

    $1.87 / $4.68in / out per M

    1.00×of list price

  • Abacus

    100% above Thinking Machines's price.

    262k context here

    $3.74 / $9.36in / out per M

    2.00×of list price

Given away with limits

1 free tiers · never ranked

These providers give Inkling away inside a quota of their own. A limit is not a price, so none of them counts as cheapest here or anywhere else on the site.

  • NVIDIA build

Frequently asked

Who sells Inkling cheapest?
DeepInfra, at $0.95 in and $4.05 out per million tokens, read on 16 Sept 2026. Hosts the open weights on its own servers, so the price is its own, not a resale. Precision not disclosed.
Is Inkling open weights?
Yes. Thinking Machines publishes the weights, so anyone may host Inkling, which is why 23 providers sell it and why their prices differ so much.
How much context does Inkling hold?
1,048,576 tokens in one call, which is about 1M, and it can call your own functions.
How does Inkling compare with the rest of the market?
At $10.29 per million tokens, Thinking Machines's price is 190% above the $3.55 median across the 337 text models we track.
More from Thinking MachinesInkling SmallSimilar contextDeepSeek V4 Flash 0731GLM-5.2DeepSeek V4 ProKimi K3Ranked in#10 in Free tiers worth using