Best AI models for writing.

334 available models considered · 42 with comparable evidence

The right writing model depends on what you want to keep: your voice, the facts in a source, a house style or the meaning of a translation. This shortlist emphasizes language and instruction following so you can choose a few candidates to try with the same brief.

For a balanced task, the current best measured match is Gemini 3.8 Flash at 85.1/100 task fit. Change the difficulty or add requirements in the finder.

Choose task difficulty

Change the range

Quick tasks

Polish a short email, shorten a paragraph or produce headline variants.

Open finder →

Default range

Balanced

Draft an article from supplied notes or edit copy against a style guide.

Open finder →

Change the range

Complex tasks

Restructure a long document while preserving its arguments, evidence and consistent voice.

Open finder →

Balanced task range

Best measured matches

Add requirements in the finder →

Task fit applies the published weights for writing. Price helps choose a provider after capability; it does not raise a weaker model into this range.

  1. Best match

    Gemini 3.8 Flash

    Google

    Task fit

    85.1 / 100

    Capability 75.8/100

    Lowest eligible listed rate

    $0.75 in · $3.75 out / M tokens

    Atlas Cloud

    Price checked 22 Sept 2026

    Measured as gemini-3.8-flash-high · high thinking. Source ↗

  2. Best value · Lowest cost

    Gemini 3.6 Flash

    Google

    Task fit

    80.2 / 100

    Capability 73.6/100

    Lowest eligible listed rate

    $0.375 in · $1.875 out / M tokens

    Kilo Gateway

    Price checked 22 Sept 2026

    Measured as gemini-3.6-flash-high · high thinking. Source ↗

  3. Task fit

    83.9 / 100

    Capability 83/100

    Lowest eligible listed rate

    $10 in · $50 out / M tokens

    Amazon Bedrock

    Price checked 22 Sept 2026

    Measured as claude-fable-5-max-effort · max effort. Source ↗

  4. Task fit

    83.5 / 100

    Capability 82.2/100

    Lowest eligible listed rate

    $9.5 in · $47.5 out / M tokens

    Jiekou

    Price checked 22 Sept 2026

    Measured as gpt-6-astra-max · max effort. Source ↗

  5. Task fit

    83.2 / 100

    Capability 78.8/100

    Lowest eligible listed rate

    $0.75 in · $3.75 out / M tokens

    DeepInfra

    Price checked 22 Sept 2026

    Measured as gemini-3.7-flash-high · high thinking. Source ↗

  6. Task fit

    82.3 / 100

    Capability 83.4/100

    Lowest eligible listed rate

    $10 in · $50 out / M tokens

    Amazon Bedrock

    Price checked 22 Sept 2026

    Measured as claude-fable-5-1-max-effort · max effort. Source ↗

  7. Task fit

    81.3 / 100

    Capability 81.6/100

    Lowest eligible listed rate

    $1.25 in · $0.15 out / M tokens

    NanoGPT

    Price checked 22 Sept 2026

    Measured as muse-spark-1.3-xhigh · xhigh effort. Source ↗

  8. Task fit

    81 / 100

    Capability 81.1/100

    Lowest eligible listed rate

    $2 in · $10 out / M tokens

    Nous Portal

    Price checked 22 Sept 2026

    Measured as gpt-5.6-sol-max · max effort. Source ↗

  9. Task fit

    80.3 / 100

    Capability 74.6/100

    Lowest eligible listed rate

    $0.75 in · $4.5 out / M tokens

    Kilo Gateway

    Price checked 22 Sept 2026

    Measured as gemini-3.5-flash-high · high thinking. Source ↗

  10. Task fit

    80.1 / 100

    Capability 80.2/100

    Lowest eligible listed rate

    $4.5455 in · $27.2727 out / M tokens

    Poe

    Price checked 22 Sept 2026

    Measured as gpt-5.5-xhigh · xhigh effort. Source ↗

How to read this guide

Language and instruction-following benchmarks are useful signals, but they do not directly measure creativity, taste, brand voice or translation quality for every language pair.

The recommendations use the same Balanced settings as the interactive finder. See the Capability methodology for sources, exact identity rules, task weights and difficulty ranges.

Other tasksgeneral useresearchcodingdata analysisautomation