Best AI model for summarising documents.

prices as of · re-ranked daily · 286 qualifying models

Reports, contracts and transcripts go in whole and come back as a paragraph. The bill is almost all input tokens, so the input price is the number that decides this job. We count only models with 128k of context or more, and 286 of the 337 we track qualify. The cheapest of them costs $0.012 per 1,000 pages at Merge Gateway. They are ranked by what the work costs, not by the price per million tokens, because a model that is cheap to prompt and dear to answer looks different once you know the shape of the job.

The cheapest model that can do this job is GLM-5.3-Flash at Merge Gateway: $0.012 per 1,000 pages, read on 16 Sept 2026.

Cheapest that qualifies

GLM-5.3-Flash

Merge Gateway

$0.012per 1,000 pages

$0.01 / $0.05 per M tokens · $0.06 at list

lowest cost per 1,000 pages

Cheapest from the maker

Command R7B Arabic

Cohere

$0.0315per 1,000 pages

$0.04 / $0.15 per M tokens · $0.0315 at list

bought from the lab that trained it

Most context

DeepSeek V4.1 Flash

TrustedRouter

$0.0304per 1,000 pages

$0.04 / $0.08 per M tokens · $0.126 at list

the longest single call on this list

What actually matters here

128k context or more
A 300-page report fits in one call at 128k tokens. Below that you split the document up and pay to re-read the overlap on every piece.
Input price decides it
Ten tokens go in for every one that comes out, so a model with a cheap input price and a dear output price is the right shape for this work.
Reasoning is optional
Summarising is not a puzzle. A reasoning model bills its thinking as output tokens, which is money spent on the half of the job that is already cheap.

Cheapest 10 for summarising documents

A page is about 600 tokens in and 60 out. A cost per 1,000 pages is that multiplied out at the row's own input and output rate. The price beside it is the standard price, three parts input to one part output per million tokens, so the two numbers answer different questions: $0.10 per million tokens is what the model costs, and $0.012 per 1,000 pages is what the work costs.

Frequently asked

What is the cheapest model for summarising documents?
GLM-5.3-Flash at Merge Gateway, $0.012 per 1,000 pages on $0.01 / $0.05 per million tokens in and out, read on 16 Sept 2026.
How is the cost per unit worked out?
A page is about 600 tokens in and 60 out. Multiply that by 1,000 and price it at each row's own input and output rate. Nothing else is counted: no cache discount, no batch rate, no free tier.
Why do only 286 models qualify?
The job needs 128k of context or more. Models that fall short are still in the index, they just cannot do this job as described.
How current are these prices?
Prices are read daily from 212 providers' public catalogs and this ranking recomputes with them. Last refresh: 16 Sept 2026.
Related jobsBest AI model for translating contentBest AI model for sorting and taggingCheapest models with 128k contextCheapest LLM API