GLM-4.7-FlashX vs Llama-3.3-70B-Instruct.

model vs model · prices as of

GLM-4.7-FlashX takes 3 of 5 rounds against Llama-3.3-70B-Instruct. Llama-3.3-70B-Instruct counters on price: $0.38 per million tokens against $0.58, for 128k context instead of 200k.

Z.aiwins 3 rounds

GLM-4.7-FlashX

$0.58/M tokens

cheapest we can explain, at Vercel AI Gateway · $0.06 per 1,000 pages

200k context · tool calling · open weights

Meta

Llama-3.3-70B-Instruct

$0.38/M tokens

cheapest we can explain, at NanoGPT · $0.0438 per 1,000 pages

128k context · tool calling · open weights

NanoGPT

List price in (lower wins)

$0.07/M

$0.10/M

List price out (lower wins)

$0.40/M

$0.32/M

Cheapest we can explain (lower wins)

$0.58/M at Vercel AI Gateway

$0.38/M at NanoGPT

Standard price (lower wins)

$0.61/M

$0.62/M

Cost per 1,000 pages (lower wins)

$0.06

$0.0438

Context

200k tokens

128k tokens

Open weights

yes

yes

Tool calling

yes

yes

Structured output

not stated

not stated

Reads images

no

no

Reasoning

yes

no

Released

19 Jan 2026

6 Dec 2024

Providers selling it

8

46

Price we measure against

the maker's own

OpenRouter

Trains on your prompts (no is better)

not stated

not stated

Providers selling both · 3

  • LLM Gateway$0.61 / $0.81
  • Merge Gateway$0.61 / $1.16
  • Vercel AI Gateway$0.58 / $2.88

GLM-4.7-FlashX then Llama-3.3-70B-Instruct, per million tokens

Go further

Add a third or fourth model, or switch the weighting, in the tool.

What moved in the last seven days

  1. Llama-3.3-70B-InstructCortecs$0.13 / $0.41 $0.75 / $0.75268% dearer ·
  2. Llama-3.3-70B-InstructEden AI$0.76 / $0.76 $0.75 / $0.750.7% cheaper ·
  3. Llama-3.3-70B-InstructNous Portal$0.08 / $0.26 $0.10 / $0.3225% dearer ·
  4. Llama-3.3-70B-InstructRequesty$0.13 / $0.40 $0.23 / $0.4038% dearer ·
  5. GLM-4.7-FlashXOfox$0.07 / $0.43 $0.07 / $0.405% cheaper ·

Frequently asked

Which is cheaper, GLM-4.7-FlashX or Llama-3.3-70B-Instruct?
Llama-3.3-70B-Instruct, at $0.38 per million tokens against $0.58 for GLM-4.7-FlashX, three parts input to one part output, read on 16 Sept 2026.
Where is each one cheapest?
GLM-4.7-FlashX: $0.58 at Vercel AI Gateway. Llama-3.3-70B-Instruct: $0.38 at NanoGPT.
Can I buy both from one provider?
Yes. 3 providers sell both: LLM Gateway, Merge Gateway, Vercel AI Gateway. The cheapest of them for the two together is LLM Gateway.
Which holds more context?
GLM-4.7-FlashX, at 200k tokens against 128k.

Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.

Related head-to-headsGemma 4 31B IT vs GLM-4.7-FlashXGemma 4 31B IT vs Llama-3.3-70B-InstructMorecompare any two modelsevery model we track