R1 0528 vs GLM-4.5.
model vs model · prices as of
R1 0528 takes 5 of 5 rounds against GLM-4.5. It is also the cheaper of the two to run, at $1.20 per million tokens against $1.40.
R1 0528
$1.20/M tokens
cheapest we can explain, at Lambda · $0.156 per 1,000 pages
164k context · tool calling · open weights
GLM-4.5
$1.40/M tokens
cheapest we can explain, at Submodel · $0.168 per 1,000 pages
131k context · tool calling · open weights
List price in (lower wins)
$0.55/M
$0.60/M
List price out (lower wins)
$2.19/M
$2.20/M
Cheapest we can explain (lower wins)
$1.20/M at Lambda
$1.40/M at Submodel
Standard price (lower wins)
$3.84/M
$4/M
Cost per 1,000 pages (lower wins)
$0.156
$0.168
Context
164k tokens
131k tokens
Open weights
yes
yes
Tool calling
yes
yes
Structured output
yes
not stated
Reads images
no
no
Reasoning
yes
yes
Released
28 May 2025
28 Jul 2025
Providers selling it
37
23
Price we measure against
the maker's own
the maker's own
Trains on your prompts (no is better)
not stated
not stated
Providers selling both · 15
- Submodel$3.65 / $1.40
NanoGPT$2.90 / $2.20
DeepInfra$3.65 / $2.80
Eden AI$4.58 / $2.80
Nous Portal$3.65 / $4
OpenRouter$3.65 / $4
TrustedRouter$3.85 / $4.22
Jiekou$4.60 / $4
R1 0528 then GLM-4.5, per million tokens · showing the 8 cheapest of 15
Go further
Add a third or fourth model, or switch the weighting, in the tool.
What moved in the last seven days
Frequently asked
- Which is cheaper, R1 0528 or GLM-4.5?
- R1 0528, at $1.20 per million tokens against $1.40 for GLM-4.5, three parts input to one part output, read on 16 Sept 2026.
- Where is each one cheapest?
- R1 0528: $1.20 at Lambda. GLM-4.5: $1.40 at Submodel.
- Can I buy both from one provider?
- Yes. 15 providers sell both: Submodel, NanoGPT, DeepInfra, Eden AI, Nous Portal, OpenRouter, and more. The cheapest of them for the two together is Submodel.
- Which holds more context?
- R1 0528, at 164k tokens against 131k.
Prices exclude tax and are per million tokens, read 16 Sept 2026 from each provider's own catalog. The standard price blends three parts input to one part output; a page is 600 tokens in and 60 out. Bars are scaled to the larger value in each row.