Best AI models for writing.
334 available models considered · 42 with comparable evidence
The right writing model depends on what you want to keep: your voice, the facts in a source, a house style or the meaning of a translation. This shortlist emphasizes language and instruction following so you can choose a few candidates to try with the same brief.
For a balanced task, the current best measured match is Gemini 3.8 Flash at 85.1/100 task fit. Change the difficulty or add requirements in the finder.
Choose task difficulty
Change the range
Quick tasks
Polish a short email, shorten a paragraph or produce headline variants.
Open finder →Default range
Balanced
Draft an article from supplied notes or edit copy against a style guide.
Open finder →Change the range
Complex tasks
Restructure a long document while preserving its arguments, evidence and consistent voice.
Open finder →Balanced task range
Best measured matches
Task fit applies the published weights for writing. Price helps choose a provider after capability; it does not raise a weaker model into this range.
Task fit
85.1 / 100
Capability 75.8/100
Lowest eligible listed rate
$0.75 in · $3.75 out / M tokens
Atlas Cloud ↗Price checked 22 Sept 2026
Measured as gemini-3.8-flash-high · high thinking. Source ↗
Task fit
80.2 / 100
Capability 73.6/100
Lowest eligible listed rate
$0.375 in · $1.875 out / M tokens
Kilo Gateway ↗Price checked 22 Sept 2026
Measured as gemini-3.6-flash-high · high thinking. Source ↗
- Claude Fable 5 ↗
Anthropic
Task fit
83.9 / 100
Capability 83/100
Lowest eligible listed rate
$10 in · $50 out / M tokens
Amazon Bedrock ↗Price checked 22 Sept 2026
Measured as claude-fable-5-max-effort · max effort. Source ↗
- GPT-6 Astra ↗
OpenAI
Task fit
83.5 / 100
Capability 82.2/100
- Gemini 3.7 Flash ↗
Google
Task fit
83.2 / 100
Capability 78.8/100
Lowest eligible listed rate
$0.75 in · $3.75 out / M tokens
DeepInfra ↗Price checked 22 Sept 2026
Measured as gemini-3.7-flash-high · high thinking. Source ↗
- Claude Fable 5.1 ↗
Anthropic
Task fit
82.3 / 100
Capability 83.4/100
Lowest eligible listed rate
$10 in · $50 out / M tokens
Amazon Bedrock ↗Price checked 22 Sept 2026
Measured as claude-fable-5-1-max-effort · max effort. Source ↗
- Muse Spark 1.3 ↗
Meta
Task fit
81.3 / 100
Capability 81.6/100
- GPT-5.6 Sol ↗
OpenAI
Task fit
81 / 100
Capability 81.1/100
Lowest eligible listed rate
$2 in · $10 out / M tokens
Nous Portal ↗Price checked 22 Sept 2026
Measured as gpt-5.6-sol-max · max effort. Source ↗
- Gemini 3.5 Flash ↗
Google
Task fit
80.3 / 100
Capability 74.6/100
Lowest eligible listed rate
$0.75 in · $4.5 out / M tokens
Kilo Gateway ↗Price checked 22 Sept 2026
Measured as gemini-3.5-flash-high · high thinking. Source ↗
- GPT-5.5 ↗
OpenAI
Task fit
80.1 / 100
Capability 80.2/100
How to read this guide
Language and instruction-following benchmarks are useful signals, but they do not directly measure creativity, taste, brand voice or translation quality for every language pair.
The recommendations use the same Balanced settings as the interactive finder. See the Capability methodology for sources, exact identity rules, task weights and difficulty ranges.