GLM-4.7-FlashX vs Llama-3.3-70B-Instruct.
model vs model · prices as of
GLM-4.7-FlashX takes 3 of 5 rounds against Llama-3.3-70B-Instruct. Llama-3.3-70B-Instruct counters on price: $0.38 per million tokens against $0.58, for 128k context instead of 200k.
GLM-4.7-FlashX
$0.58/M tokens
cheapest we can explain, at Vercel AI Gateway · $0.06 per 1,000 pages
200k context · tool calling · open weights
Llama-3.3-70B-Instruct
$0.38/M tokens
cheapest we can explain, at NanoGPT · $0.0438 per 1,000 pages
128k context · tool calling · open weights
List price in (lower wins)
$0.07/M
$0.10/M
List price out (lower wins)
$0.40/M
$0.32/M
Cheapest we can explain (lower wins)
$0.58/M at Vercel AI Gateway
$0.38/M at NanoGPT
Standard price (lower wins)
$0.61/M
$0.62/M
Cost per 1,000 pages (lower wins)
$0.06
$0.0438
Context
200k tokens
128k tokens
Open weights
yes
yes
Tool calling
yes
yes
Structured output
not stated
not stated
Reads images
no
no
Reasoning
yes
no
Released
19 Jan 2026
6 Dec 2024
Providers selling it
8
46
Price we measure against
the maker's own
OpenRouter
Trains on your prompts (no is better)
not stated
not stated
Providers selling both · 3
LLM Gateway$0.61 / $0.81
Merge Gateway$0.61 / $1.16
Vercel AI Gateway$0.58 / $2.88
GLM-4.7-FlashX then Llama-3.3-70B-Instruct, per million tokens
Go further
Add a third or fourth model, or switch the weighting, in the tool.
What moved in the last seven days
- Llama-3.3-70B-Instruct
Cortecs$0.13 / $0.41 $0.75 / $0.75268% dearer · - Llama-3.3-70B-Instruct
Eden AI$0.76 / $0.76 $0.75 / $0.750.7% cheaper · - Llama-3.3-70B-Instruct
Nous Portal$0.08 / $0.26 $0.10 / $0.3225% dearer · - Llama-3.3-70B-Instruct
Requesty$0.13 / $0.40 $0.23 / $0.4038% dearer · - GLM-4.7-FlashX
Ofox$0.07 / $0.43 $0.07 / $0.405% cheaper ·
Frequently asked
- Which is cheaper, GLM-4.7-FlashX or Llama-3.3-70B-Instruct?
- Llama-3.3-70B-Instruct, at $0.38 per million tokens against $0.58 for GLM-4.7-FlashX, three parts input to one part output, read on 16 Sept 2026.
- Where is each one cheapest?
- GLM-4.7-FlashX: $0.58 at Vercel AI Gateway. Llama-3.3-70B-Instruct: $0.38 at NanoGPT.
- Can I buy both from one provider?
- Yes. 3 providers sell both: LLM Gateway, Merge Gateway, Vercel AI Gateway. The cheapest of them for the two together is LLM Gateway.
- Which holds more context?
- GLM-4.7-FlashX, at 200k tokens against 128k.
Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.