DeepInfra vs Novita.

model prices compared · prices as of · 64 models both sell

DeepInfra is cheaper on 48 of the 64 models both sell, Novita on 9, and 7 cost the same at both. Both are the same kind of provider, hosts running the same models, so the price is most of what separates them. The widest gap is DeepSeek V4 Flash 0731, $0.45 per million tokens at DeepInfra against $2.46.

DeepInfra is cheaper on 48 of the 64 models both sell, and the cheapest thing either one offers is Llama 3.2 3B Instruct at $0.08 per million tokens, read 16 Sept 2026.

Serves open-weight models on its own GPUs, often at fp8 or fp4, and publishes the precision per model.

cheapest here: Llama 3.2 3B Instruct at $0.08/M tokens

Serves open-weight models on its own GPUs; publishes list and promotional prices side by side.

cheapest here: Llama 3.1 8B Instruct at $0.11/M tokens

Kind of providerHostHost
Models we track10692
Cheaper of the two48 of 649 of 64
Trains on your promptsdoes not saydoes not say
Keeps your requestsnot statednot stated
How you paynot statednot stated
Smallest top-upnot statednot stated
Free tiernone publishednone published
Where it runsnot statednot stated
Prices read16 Sept 202616 Sept 2026

Every model both sell, at each one's price

Most widely sold first. Each price is the cheapest that provider charges for the model under the standard blend, three parts input to one part output per million tokens. A free tier is a limit rather than a price, and a price below the maker's own with nothing published to account for it is left out, so neither can be marked as the cheaper of the two.

ModelDeepInfraNovita
DeepSeek V4 Flash 0731DeepSeek · 1M context$0.45$0.09 / $0.18 in / out$2.46$0.41 / $1.23 in / out
GLM-5.2Z.ai · 1M context$4.65$0.75 / $2.40 in / out$3.99$0.65 / $2.04 in / out
DeepSeek V4 ProDeepSeek · 1M context$6.50$1.30 / $2.60 in / out$5.94$0.99 / $2.97 in / out
Kimi K2.6Moonshot AI · 262k context$5.75$0.75 / $3.50 in / out$5.80$0.80 / $3.40 in / out
Kimi K3Moonshot AI · 1M context$22.80$2.85 / $14.25 in / out$24$3 / $15 in / out
GLM-5.3-FlashZ.ai · 1M context$0.95$0.15 / $0.50 in / out$0.84$0.13 / $0.44 in / out
GLM-5.3Z.ai · 1M context$7.60$1.20 / $4 in / out$5.59$0.91 / $2.86 in / out
Kimi K2.7 CodeMoonshot AI · 262k context$5.44$0.68 / $3.40 in / out$6.58$0.91 / $3.84 in / out
gpt-oss-120bOpenAI · 131k context$0.28$0.04 / $0.17 in / out$0.40$0.05 / $0.25 in / out
GLM-5.1Z.ai · 200k context$6.65$1.05 / $3.50 in / out$8.54$1.38 / $4.40 in / out
MiniMax-M3MiniMax · 1M context$1.94$0.28 / $1.10 in / out$2.10$0.30 / $1.20 in / out
DeepSeek V3.2DeepSeek · 164k context$1.16$0.26 / $0.38 in / out$1.21$0.27 / $0.40 in / out
DeepSeek V4.1 FlashDeepSeek · 1M context$1.20$0.20 / $0.60 in / out$2.10$0.30 / $1.20 in / out
Kimi K2.5Moonshot AI · 262k context$3.60$0.45 / $2.25 in / out$4.80$0.60 / $3 in / out
GLM-5Z.ai · 205k context$3.88$0.60 / $2.08 in / out$6.20$1 / $3.20 in / out
Qwen3.8 27BAlibaba Cloud · 1M context$2.33$0.15 / $1.88 in / out$4.26$0.42 / $3 in / out
MiniMax-M2.7MiniMax · 205k context$1.75$0.25 / $1 in / out$1.89$0.27 / $1.08 in / out
MiniMax-M2.5MiniMax · 205k context$1.60$0.15 / $1.15 in / out$2.10$0.30 / $1.20 in / out
Llama-3.3-70B-InstructMeta · 128k context$0.62$0.10 / $0.32 in / out$0.81$0.14 / $0.40 in / out
Qwen3.5 397B-A17BAlibaba Cloud · 262k context$4.35$0.45 / $3 in / out$5.40$0.60 / $3.60 in / out
gpt-oss-20bOpenAI · 131k context$0.23$0.03 / $0.14 in / out$0.27$0.04 / $0.15 in / out
GLM-4.7Z.ai · 205k context$2.95$0.40 / $1.75 in / out$3.60$0.54 / $1.98 in / out
Gemma 4 31B ITGoogle · 262k context$0.61$0.09 / $0.34 in / out$0.82$0.14 / $0.40 in / out
Qwen3.6 35B-A3BAlibaba Cloud · 262k context$1.25$0.10 / $0.95 in / out$2.23$0.25 / $1.49 in / out
Qwen3 235B-A22BAlibaba Cloud · 131k context$0.82$0.09 / $0.55 in / out$0.85$0.09 / $0.58 in / out
R1 0528DeepSeek · 164k context$3.65$0.50 / $2.15 in / out$4.60$0.70 / $2.50 in / out
GLM-4.6Z.ai · 205k context$3.50$0.50 / $2 in / out$3.85$0.55 / $2.20 in / out
MiMo-V2.5-ProXiaomi · 1M context$6$1 / $3 in / out$2.61$0.52 / $1.04 in / out
Qwen3 32BAlibaba Cloud · 131k context$0.52$0.08 / $0.28 in / out$0.75$0.10 / $0.45 in / out
MiMo-V2.5Xiaomi · 1M context$0.70$0.14 / $0.28 in / out$0.84$0.17 / $0.34 in / out
Qwen3.6 27BAlibaba Cloud · 262k context$4.16$0.32 / $3.20 in / out$5.40$0.60 / $3.60 in / out
Gemma 4 26B A4B ITGoogle · 262k context$0.55$0.07 / $0.34 in / out$0.79$0.13 / $0.40 in / out
Qwen3 30B A3B Instruct 2507Alibaba Cloud · 262k context$0.86$0.12 / $0.50 in / out$0.72$0.09 / $0.45 in / out
Hy3Tencent · 262k context$1$0.14 / $0.58 in / out$1$0.14 / $0.58 in / out
Kimi K2 0905Moonshot AI · 262k context$3.50$0.50 / $2 in / out$4.30$0.60 / $2.50 in / out
Qwen3.5 35B-A3BAlibaba Cloud · 262k context$1.42$0.14 / $1 in / out$2.75$0.25 / $2 in / out
DeepSeek V4 Flash Vision ExpDeepSeek · 1M context$2.64$0.44 / $1.32 in / out$2.64$0.44 / $1.32 in / out
Mistral NemoMistral AI · 128k context$0.09$0.02 / $0.03 in / out$0.29$0.04 / $0.17 in / out
Qwen3-Next 80B-A3B InstructAlibaba Cloud · 131k context$1.37$0.09 / $1.10 in / out$1.95$0.15 / $1.50 in / out
Llama 3.1 8B InstructMeta · 131k context$0.10$0.02 / $0.04 in / out$0.11$0.02 / $0.05 in / out

showing the 40 most widely sold of 64 · every row in the open dataset

Frequently asked

Is DeepInfra cheaper than Novita?
On the 64 models both sell, DeepInfra is cheaper on 48, Novita on 9, and 7 cost the same at both. The widest gap is DeepSeek V4 Flash 0731: $0.45 per million tokens against $2.46. Prices are blended three parts input to one part output and read on 16 Sept 2026.
Which models do DeepInfra and Novita both sell?
64 of the models we track, led by DeepSeek V4 Flash 0731, GLM-5.2, DeepSeek V4 Pro. DeepInfra sells 106 in all and Novita 92, so most of each catalog has no counterpart at the other.
Does DeepInfra or Novita train on what you send it?
DeepInfra does not say train on what you send it, and not stated. Novita does not say, and not stated. A provider that may train on prompts is often how a price below the maker's own is paid for.
How current are these prices?
Both catalogs are read every day. DeepInfra was last read 16 Sept 2026 and Novita 16 Sept 2026; the whole index was refreshed 16 Sept 2026. Prices exclude tax, and a price below the maker's own with nothing published to account for it is left out of this table entirely.
Provider pageseverything DeepInfra sellseverything Novita sellsMore matchupsDeepInfra vs Fireworks AIDeepInfra vs GroqDeepInfra vs Nebius Token FactoryDeepInfra vs OpenRouterDeepInfra vs Together AINebius Token Factory vs Novita