ArticlesWriting tools

The best AI writing assistants in 2026, sorted by the problem they actually solve

I merged deep research from three AI systems, checked the claims against primary sources, and dropped what I could not confirm. 453 professionals, 33 experiments, 2,582 writers, and a warning about AI-written apologies.

By Samet Durgun · Co-founder of Subtext · 14 min read

Every list of AI writing assistants I found this year was either an affiliate page or a vendor blog ranking itself first. I wanted something I could trust, so I built it myself. I ran the same research brief through three deep research systems (ChatGPT, Claude, and Gemini), merged the reports, then checked the claims that mattered against primary sources. Some numbers changed along the way. Two studies never turned up anywhere, so they are gone. Everything that survived links to where it came from.

One disclosure before we start. I build Subtext, an app about how messages land with the person reading them. That makes me biased toward the interpersonal side of this topic. It also means I read communication research for fun, which will become obvious in the second half of this post.

The short answer

If you write mostly in one language and want polish while you type, Grammarly remains the strongest all-round editor. If your problem is finding better wording for a sentence you already have, test Wordtune and DeepL Write. If you move between German and English like I do, put DeepL Write, LanguageTool, and Grammarly through the same trial week and keep whichever annoys you least. For difficult messages, no grammar checker will help. Either brief a chatbot properly or use Subtext, which I co-founded and which asks for the context and reads your draft before it changes a word. For email volume, the Gemini or Copilot already inside your inbox covers most of it, and Superhuman is the upgrade once email is the job.

And the finding that changed how I use these tools myself: several studies show people trust you less when they suspect AI wrote your message, with the harshest penalty on the warm, personal ones. That research is further down.

No study proves one tool beats another

I went looking for an independent benchmark that ranks Grammarly against Wordtune against QuillBot on real writing. It does not exist. The closest thing is a 2026 meta-analysis in Assessing Writing1 covering 33 experiments on algorithm-based writing feedback. Feedback helped overall (g = 0.36, a small to medium effect), second-language writers gained the most, and the differences between individual tools were not statistically significant. The detail that made me smile: feedback from Grammarly and Pigai measurably improved student writing, and feedback from ChatGPT alone did not reach significance.

Zapier tested more than 50 writing tools2 and reached a compatible conclusion from the product side: raw text generation has become similar enough across tools that workflow, control, and integrations decide which one earns a place in your day. So when a listicle tells you tool A is 94 percent accurate and tool B is 87, treat it as decoration. Nobody has published a methodology behind numbers like that.

Writing assistant now means four different products

The category has split. Editors like Grammarly, LanguageTool, DeepL Write, and ProWritingAid react to text you already wrote. Wordtune and QuillBot are rewriters, built to generate alternatives for a sentence or a paragraph. ChatGPT, Claude, and Gemini reason about your situation first and write second. And platforms like Jasper, Writer, and Anyword exist to produce marketing or enterprise text at scale. LanguageTool draws a version of this line on its own site, separating tools that improve your text from tools that generate it for you. Email clients with AI built in form a fifth group, covered separately below.

The editors, ranked by the problems they solve

Grammarly stays the default for a reason. It works in the browser, in Gmail, in Docs, on your phone, and it reacts while you type instead of waiting for a prompt. Its tone detector3 labels how a draft may come across, from confident to worried, and the rewrite features adjust formality on request. The company reports more than 40 million daily users. Pro costs about 12 dollars a month on the annual plan, 30 on monthly, with 2,000 AI prompts included. Two caveats. The tone labels are product claims, and no independent study validates them against how real recipients react. Its rewrites also tend to sand off personality, which the homogenisation research below should make you take seriously. Worth knowing: Grammarly bought the email client Superhuman in 2025 and took its name for the parent company, so expect deeper email bundling.

DeepL Write was missing from my first research pass and deserves a top slot, especially in Europe. It corrects grammar, offers alternative phrasings for whole sentences, and switches between styles like business, academic, and casual, plus tones like friendly, confident, and diplomatic. Coverage spans English, German, French, Italian, Portuguese, and Spanish. I write English and German daily, sometimes Turkish, and for the first two DeepL Write comes closest to how a bilingual colleague would rephrase things. Write Pro runs around 7.50 dollars a month. The gap for me: no Turkish yet.

LanguageTool competes on languages rather than flash. It checks grammar and style in more than 30 languages, with the deepest coverage for European ones including German, and offers an open-source core plus self-hosting for teams that care where their text goes. Premium lands near five dollars a month on longer plans. The output feels quieter than Grammarly, fewer suggestions, less rewriting, which some people will prefer.

Wordtune solves one moment well: you know what you mean and the sentence refuses to say it. It generates several rewrites at different formality levels, shortens or expands, and continues a thought when you stall. The free tier gives ten rewrites a day, and paid plans have floated between roughly 7 and 10 dollars a month depending on promotions. For emails and short messages, this stays the most focused pure rewriting tool I tested.

QuillBot is the paraphrasing machine. Ten modes, from fluency to academic to a humanizer, plus a slider controlling how far the vocabulary drifts from your original, for around 8 dollars a month on annual billing. It transforms tone on command. Whether the new version will land with the specific person reading it is a question no paraphraser answers, and that gap matters more for messages than for essays.

ProWritingAid went deep instead of wide. More than 25 analysis reports covering pacing, repetition, readability, and cliches, plus manuscript-level analysis for fiction. The design decision I respect most: it separates assistive AI, which analyses without rewriting, from generative AI, which produces new text. Around 10 dollars a month, with a lifetime license near 399. For a novel, I would pick it over Grammarly. For a WhatsApp message, no.

ChatGPT, Claude, and Gemini play a different game

A grammar checker sees the sentence in front of it. A chatbot can hold the whole situation: who the recipient is, what happened before, what you want to achieve, what you fear the message sounds like. For anything delicate, that context beats correction.

Each has a distinct edge. ChatGPT added canvas-style drafts you edit directly inside the conversation, which turns it into a workable writing surface rather than a chat log. Claude builds custom styles from writing samples you upload, so it can learn how you sound instead of offering a formal or casual toggle, and its Projects keep long documents and their context together. Gemini moved into the Google stack: since spring 2026 it drafts documents in Docs4 using material from your Drive, Gmail, and Chat, and its Match writing style feature edits a document toward one consistent voice.

The catch comes from a 2026 study in Computers and Composition5. Researchers asked eight generative AI platforms, including versions of Claude, Gemini, Copilot, and ChatGPT, to edit academic writing samples, then compared the output with human editing. The models caught fewer sentence-level errors than the human baseline, over-flagged some error types, and missed others. These systems rewrite well and proofread worse than their confidence suggests. I run every AI-touched draft through a dedicated checker afterward, and that study is why.

For email, start with what you already pay for

Gemini in Gmail drafts and rewrites replies with thread context and has started personalising suggestions using your past mail. Copilot does the equivalent in Outlook, with drafting, tone coaching, and thread summaries, though it needs a Microsoft 365 Copilot license. On iPhone and Mac, Apple Intelligence rewrites texts and emails in different tones, free and on-device. Reviewers6 describe it as competent and conservative, and I agree with both words.

Paid upgrades exist for people who live in email. Superhuman pre-writes replies in your voice and adapts tone per recipient based on your history with that person, at around 30 to 40 dollars a month with the voice features on the Business tier. Superhuman says that in testing, 40 percent of its auto drafts got sent within a day and most of those went out unedited. Both numbers come from Superhuman, so treat them as marketing until someone independent measures them. Shortwave does cheaper voice-matching for Gmail only, though its pricing gets reported so inconsistently across sources that I would confirm on their site before subscribing. Missive fits teams sharing an inbox and lets you pick which model powers it. And if you remember Flowrite, it was acquired in January 20257 and folded into MailMaestro.

The specialists, briefly

Fiction writers get Sudowrite, built around a prose-tuned model called Muse and a Story Bible that keeps characters and plot consistent across chapters, from about 10 dollars a month. Academic writers get a whole sub-ecosystem: SciSpace for discovery and drafting, Elicit for systematic reviews, Consensus for quick evidence checks, Paperpal for submission-ready manuscripts. The comparison numbers floating around for these mostly come from the vendors’ own benchmark pages, so read them the way you read a pitch deck. Marketing teams get Jasper, which turns brand voice and style rules into governed output at scale from 39 dollars a month, and Anyword, which scores copy with a predictive performance model; the accuracy figure Anyword cites for that score comes from Anyword. Enterprises get Writer, which exists to keep a company’s terminology, style, and compliance consistent across thousands of employees.

What the research says AI does to your writing

The strongest evidence covers speed and polish

The best causal study remains Noy and Zhang’s randomised experiment in Science8. They assigned occupation-specific writing tasks to 453 college-educated professionals and let half use ChatGPT. Average completion time dropped 40 percent, rated quality rose 18 percent, and the weakest writers improved the most. That was GPT-3.5 in early 2023, and current models are stronger.

Above the sentence level, the evidence thins out

A 2026 systematic review in the Journal of English for Academic Purposes9 synthesised 25 empirical studies and found a division of labour: chatbots handle ideation and early drafting, automated feedback tools handle revision. Reported gains cluster around surface quality and writing process, with weaker and more mixed evidence for argumentation and structure. The Assessing Writing meta-analysis points the same direction, with a twist: where deep-level improvements appeared, they held up better over time than the surface-level ones did.

Everyone starts sounding the same

Doshi and Hauser ran an experiment on short stories10 for Science Advances. Writers with access to AI ideas produced stories rated more creative and better written, especially the less creative writers. The same stories were also measurably more similar to each other. Individually you improve, collectively everyone converges. Related work on the AI Ghostwriter Effect11 found people willingly sign their name under AI text they do not feel ownership of, which is its own quiet problem.

Autocomplete can move your opinions without you noticing

In March 2026, Science Advances published two preregistered experiments with 2,582 participants12 who wrote about societal issues with an AI assistant feeding them subtly biased autocomplete suggestions. Their stated opinions afterward had moved toward the AI’s position. Most never noticed. Warning them beforehand did not help12, and neither did debriefing afterward. The effect ran stronger through interactive suggestions than when the same arguments appeared as static text. Since autocomplete now lives in Gmail, keyboards, and half the tools above, this study earned a permanent spot in my head.

And a pointer at where this goes next: alignment research is moving toward inferring your personal style rules from samples of your writing13 in place of one-off prompts, which suggests where the personalization race goes next.

How people judge you when they suspect AI wrote it

The SEO listicles skip this section, and for messages it matters more than any feature.

Cornell researchers put people in conversations with AI smart replies and published the results in Scientific Reports14. The replies made conversations faster and the language more positive. Yet when someone merely suspected their partner used AI, they rated that partner as less cooperative and felt less close to them14. The suspicion did the damage, whether or not AI was involved.

A CHI 2022 study15 tested condolence and support emails. Trust in the writer dropped when AI involvement was disclosed, and participants described AI-written emotional messages as fake. Earlier work had already named the pattern: the Replicant Effect16, where profiles suspected of machine authorship get trusted less.

The 2025 study I cite most often comes from USC and the University of Florida17. More than a thousand working professionals evaluated workplace messages written with different levels of AI help. Supervisors using heavy AI assistance were rated sincere by roughly 40 to 52 percent of employees, against 83 percent for low-assistance messages, and perceived professionalism fell in parallel17. People also judged their own AI use far more leniently than their boss’s. The authors’ practical read matches mine: minor AI edits to professional email pass without cost, and heavy generation is where respect erodes.

A 2025 paper presented at AIES18 adds the mechanism. Messages labeled AI-assisted read as weaker signals of character, in both directions: an AI-assisted apology seems less warm, an AI-assisted blame seems less cold. Unassisted writing works as a costly signal, proof you spent effort on the other person.

For balance: a 2025 iScience study19 found AI mediation can speed up trust-building in some interactions. The penalty concentrates where a message is supposed to carry effort and feeling, and softens for transactional exchanges.

So my personal rule, backed by everything above: full AI for logistics, light AI for professional email, and for apologies, condolences, and anything relationship-defining, write it yourself and use AI only to check the grammar. Wanting help with hard messages without outsourcing the humanity is exactly the tension I care about with Subtext, and this research keeps confirming it is real.

For the message you keep rewriting

Every category above answers some version of “make this better”. Editors correct what you wrote, rewriters generate alternatives, chatbots reason about your situation once you explain it, and marketing platforms produce copy at volume. For a difficult personal message the useful question is the other one: what does this say right now, before I send it.

The research in the section above is what convinced me that question outranks the rewrite. If heavy generation costs you sincerity in the reader’s eyes, and unassisted writing reads as a costly signal of effort, then what you want is a tool that hands the draft back to you rather than replacing it. On my own rule from a moment ago, write it yourself and use AI only to check, this is the checking step.

That is the job Subtext is built for, and I co-founded it, so treat this as an argument rather than a review. It takes the thread rather than the sentence, names what your draft is likely to signal to the person receiving it, and points at the wording doing the work, so what you send stays in your own words. A chatbot can reach the same place if you paste the context and ask the right question. The difference is which one of you has to remember to do that at 11pm.

Editors Rewriters Chatbots Subtext
Grammar and style correction Yes Partly Yes Yes
Rewrites in a chosen tone Yes Yes Yes Yes
Tells you how the draft is likely to land Tone labels No If you ask Yes
Reads the thread, not just the message No No If you paste it Yes
Asks who is receiving it No No No Yes
Built for Documents and general writing Sentences and paragraphs Anything, if you prompt it well Messages between people

Editors here are Grammarly, LanguageTool and DeepL Write, rewriters are Wordtune and QuillBot, chatbots are ChatGPT, Claude and Gemini.

The setup I would run

Two subscriptions cover most writing:

  1. One editor that lives where you type. Grammarly if you write mainly English, DeepL Write or LanguageTool if you work across languages.
  2. One chatbot for context-heavy work, picked from whichever ecosystem you already inhabit.

Then the situational picks: Subtext for the message you keep rewriting and still have not sent, Wordtune when a sentence resists you, ProWritingAid for a manuscript, Superhuman once you send dozens of emails a day, Jasper or Writer once a whole team needs one voice, Sudowrite for fiction.

Prices in this post were checked in August 2026, and these companies change plans constantly, so trust their pricing pages over mine.

The pattern across everything above: base capability has converged, and what separates tools now sits around the model. Where it fits your workflow, and what it does to the person reading. That second question is the one I keep pulling on, in this research and in what I build.


If you spot a study I missed or a claim that deserves a harder look, write me. I would rather correct this post than defend it.

Samet Durgun is the co-founder of Subtext, an app that catches the emotional tone of your messages and rewrites them in your own voice. He’s based in Berlin.


Sources

Every link above goes to the primary source where one exists. Research and reporting are numbered in order of appearance; the product pages below are linked inline where they come up and are listed here for reference. Prices and features change, so verify before buying.

  1. Can algorithm-based feedback help students to write better? A meta-analysis, Assessing Writing, 2026
  2. Zapier, best AI writing generators, tested 50+ tools
  3. TechCrunch on Grammarly’s tone detector
  4. Google Workspace Updates: Gemini in Docs, April 2026
  5. Is genAI a good editor of academic writing?, Computers and Composition, 2026
  6. Six Colors review of Apple Intelligence
  7. Maestro Labs acquires Flowrite, January 2025
  8. Noy & Zhang, Experimental evidence on the productivity effects of generative AI, Science, 2023
  9. A systematic review of AI-assisted academic writing, Journal of English for Academic Purposes, 2026
  10. Doshi & Hauser, Generative AI enhances individual creativity but reduces the collective diversity of novel content, Science Advances, 2024
  11. Draxler et al., The AI Ghostwriter Effect, ACM TOCHI, 2024
  12. Williams-Ceci et al., Biased AI writing assistants shift users’ attitudes on societal issues, Science Advances, 2026 and the Cornell summary
  13. Aligning LLMs by predicting preferences from user writing samples, arXiv, 2025
  14. Hohenstein et al., Artificial intelligence in communication impacts language and social relationships, Scientific Reports, 2023 and the Cornell Chronicle summary
  15. Liu et al., Will AI console me when I lose my pet?, CHI 2022
  16. Jakesch et al., AI-mediated communication and the Replicant Effect, CHI 2019
  17. Professionalism and trustworthiness in AI-assisted workplace writing, International Journal of Business Communication, 2025 and the PsyPost summary
  18. Khadpe et al., on AI assistance and judgments of character, AIES 2025
  19. AI-mediated communication and trust-building, iScience, 2025
  20. Zapier, best AI email assistants

Vendor documentation