ArticlesMessages

Does this text sound rude?

You hear your own intended tone while you write and cannot switch it off. That makes you reliably overconfident about irony, teasing and the line between firm and cold, and reasonably accurate about plain warmth. What actually helps, ranked by evidence.

By Samet Durgun · Co-founder of Subtext · 9 min read

You have read it back six times and it sounds fine. It also sounded fine the first time, which is the problem, because reading it back is the one test that cannot fail. I co-founded Subtext because this question, asked of a message the person has already reread, is the most common thing anyone brings to it.

The honest answer is that you cannot tell from the inside, and the reason is well characterised. You silently hear your own intended tone while you write, and you cannot subtract it when predicting how someone else will read the words without it. That is the mechanism. But the popular version overstates the damage. You are bad at conveying irony, teasing, and the difference between firm and cold. You are not demonstrably bad at conveying ordinary warmth to people who already know you. Both halves matter, and so does knowing which half your message falls into.

The study behind everything, with the numbers labelled

Kruger, Epley, Parker and Ng ran five experiments on this and the paper is cited everywhere, usually with figures that do not say which study they came from1. Here they are, labelled, because the unlabelled versions have caused real confusion.

In the first study, twelve students in six pairs wrote serious and sarcastic statements and emailed them. Senders predicted 97 per cent would be decoded correctly. Actual accuracy was 84. Reliable overconfidence on a tiny sample, and note that actual accuracy was still high. The finding is “worse than they thought”, not “bad”.

In the third and largest study, 154 pairs were tested on sarcasm, seriousness, anger and sadness. Senders predicted 88.8 per cent accuracy and readers achieved 70.4. Friends were no better than strangers. And face-to-face was no better than voice-only, which points at missing intonation rather than missing gesture as the operative loss.

In the second study, email readers detected sarcasm at a rate the authors describe as indistinguishable from chance, while people hearing the same lines aloud got roughly three quarters right. You will see that study quoted as 78 per cent predicted against 56 achieved. Those figures trace to a magazine summary rather than the paper, which uses the phrasing above.

The mechanism test is the fourth study, and it is the one that justifies everything practical in this article. Participants read their own statements aloud either in the intended tone or deliberately in the wrong one. Reading them the wrong way erased the overconfidence entirely. Your own inner voice is the culprit, and this is the experiment that shows it.

Two things to carry. The paper has never been directly replicated at high power. It fell outside the sampling frame of the big replication projects and appears in none of them. It is famous, theoretically coherent and internally consistent, and it rests on small mid-2000s undergraduate samples. And the mechanism, egocentric anchoring, has theoretical parents in the curse of knowledge and the illusion of transparency, neither of which has a gold-standard direct replication either.

The finding that narrows the thesis

Pollmann and Roos tested the popular assumption directly, using real messages and asking both people2. Receivers rated the warmth of a message they had actually received; the sender then rated what they had intended. Across 171 matched pairs for text messages and 61 for work emails, the ratings aligned closely, and six moderators including message length and emoji changed nothing.

Read the scope carefully, because it is what makes the two findings compatible rather than contradictory.

Kruger and colleagues Pollmann and Roos
Materials Invented sentences on assigned topics Real messages people actually sent
Relationship Strangers or light acquaintances Existing relationships
What was judged Which specific tone was intended Warm to cold, one dimension
Difficulty Deliberately ambiguous, irony-heavy Whatever people naturally write

The synthesis: people are good at conveying whether they feel good or bad, and bad at conveying irony, teasing, and the line between firm and cold, especially with people who do not know them well.

Pollmann and Roos measured valence, a single dimension. “Was this warm or cold” is a different question from “did you catch that I was joking” or “did this read as passive-aggressive”, and their null result does not licence a general claim that written tone is understood. It is also a single study in a companion open-access journal, not yet independently replicated.

So when you ask whether your message sounds rude, the first thing to check is which kind of message it is. Plain and warm to someone who knows you: probably fine. Dry, ironic, or firm to someone who does not: this is exactly where the gap lives, and it is the only kind of message Subtext bothers flagging.

Why your intent feels audible

Most people hear a voice when they read. In a general-population survey of 570 people, 80.7 per cent reported sometimes or always hearing an inner voice during silent reading, and most described it as having gender, accent, pitch and emotional tone3. The 19.3 per cent who reported no inner reading voice is a useful complication, since it means this mechanism is probably not universal.

Linguists call the related phenomenon implicit prosody: silent reading projects a default melody onto text, and that melody influences how you parse it4. Brain measures back it up, with commas and implicit phrase boundaries producing responses that resemble the response to spoken boundaries.

The honest limit: no single study measures inner-voice prosody and then links it to a specific misjudgement of your own written tone. What exists is two well-supported literatures meeting at a plausible junction, plus Kruger’s fourth study, which manipulates the sender’s vocalisation and gets the predicted result.

This is also the precise gap Subtext sits in. It reads the message with no inner voice at all, which is the one thing you cannot do to your own draft, and it reports which readings the words support rather than the one you meant.

Why rereading does not work, and typos survive

The same mechanism explains the typo you missed five times.

Daneman and Stainton had people write an essay and then proofread it after either twenty minutes or two weeks, alongside a familiar and an unfamiliar essay by someone else5. They caught fewer errors in their own writing than in unfamiliar writing. And the two-week delay reduced the disadvantage, which shows the problem is extreme familiarity rather than authorship as such.

One correction worth making, because the wrong explanation circulates widely. The claim that the generation effect, meaning better memory for things you produced yourself, is why you miss your own typos has been directly tested and not confirmed. Two eye-tracking experiments found no self-generation effect in proofreading6. The robust account is familiarity and predictability: you know what the text is meant to say, so you see that rather than what is there.

Your expertise with your own message actively hurts. That is why the copy editor with no idea what you meant does better.

What actually helps, ranked

Wait. The best-evidenced fix. Daneman and Stainton is direct experimental evidence that a delay improves error detection in your own writing5. For tone the delay lets the inner voice fade. Two weeks is not practical for a text message. Overnight usually is.

Read it aloud in the wrong tone. Kruger’s fourth study is the evidence1. Deliberately not performing your intended tone dissolves the overconfidence. Reading it aloud in the tone you meant may do little, because that reinforces your own phenomenology. Read it flat, or read it as if you were annoyed, and see what survives.

Read it aloud for errors. Now directly supported. Cushing and Bodner compared proofreading aloud, silently, and in a deliberately hard-to-read font7. Aloud improved error detection. The disfluent font either impaired it or did nothing. The interesting part is that participants did not predict the aloud benefit, so this is useful advice people will not adopt on intuition. Sample sizes for both experiments could not be verified from the accessible materials, so no numbers.

Get an outside reader. Well supported for clarity and errors, since a fresh reader lacks your privileged knowledge, which is the whole advantage the proofreading studies document. One caveat for tone specifically: any negativity bias belongs to the actual recipient, so a neutral third party may not reproduce the reaction of one specific anxious or subordinate reader.

Change the format. Changing the typescript restores error salience8. This is not the same as using a hard-to-read font, which the Cushing and Bodner work found unhelpful.

Say the thing explicitly. Directly tested, and it addresses the best-evidenced asymmetry in this whole literature, which is about urgency rather than rudeness. Giurge and Bohns ran eight preregistered experiments with 4,004 working adults and found receivers overestimate how fast senders expect replies to non-urgent out-of-hours email, and pay for it in stress9. A brief note from the sender saying there is no rush reduced the bias. Your convenient Sunday-night email reads as a demand, and the fix is four words long.

Text-to-speech. Principled rather than proven. It supplies a genuinely external voice with none of your intended prosody, which is what the fourth Kruger study suggests should help. I found no direct experiment isolating it for tone.

What Subtext does with this

It is the outside reader, minus the inner voice, plus the specific check for the messages where the gap actually lives. It flags irony that will not travel, since figuring out whether a message is sarcastic from the words alone has its own research and its own limits, a firm sentence that reads cold, and a warm message that has been edited into a flat one. It does not flag plain warm messages to people who know you, because the evidence says those mostly land.

This is the same gap our tone-detection breakdown covers from the product side, if you want to see what that reading actually looks like on a real message. Subtext is the outside reader without an inner voice. Free to start, on your phone or in the browser.

Sources

Numbered in the order they appear above. Where a figure could not be verified against the primary source, the text says so.

  1. Kruger, J., Epley, N., Parker, J., and Ng, Z.-W. (2005). Egocentrism over e-mail: Can we communicate as well as we think? Journal of Personality and Social Psychology, 89(6), 925 to 936. Funded by a University of Illinois Board of Trustees research grant and NSF grant SES-0241544. Study 1: 12 students, six dyads, 97 against 84. Study 3: 154 pairs, 88.8 against 70.4. Study 4: 54 students. Never directly replicated at high power. https://psycnet.apa.org/record/2005-15488-002
  2. Pollmann, M. M. H., and Roos, C. A. (2025). “I get u”. People correctly interpret the tone of text messages and emails. Computers in Human Behavior Reports, 18, article 100689. Measures valence only. Companion open-access title; not yet independently replicated. Funding unverified. https://doi.org/10.1016/j.chbr.2025.100689
  3. Vilhauer, R. P. (2017). Characteristics of inner reading voices. Scandinavian Journal of Psychology, 58(4), 269 to 274. General-population survey, n = 570. Self-report about reading in general. Funding unverified. https://onlinelibrary.wiley.com/doi/10.1111/sjop.12368
  4. Alderson-Day, B., and Fernyhough, C. (2015). Inner speech: Development, cognitive functions, phenomenology, and neurobiology. Psychological Bulletin, 141(5), 931 to 965. The major review of inner speech. https://psycnet.apa.org/record/2015-24380-001
  5. Daneman, M., and Stainton, M. (1993). The generation effect in reading and proofreading: Is it easier or harder to detect errors in one’s own writing? Reading and Writing, 5(3), 297 to 313. https://link.springer.com/article/10.1007/BF01027393
  6. Burgoyne, A. P., Saba-Sadiya, S., Harris, L. J., Becker, M. W., Brascamp, J. W., and Hambrick, D. Z. (2023). Revisiting the self-generation effect in proofreading. Psychological Research, 87, 800 to 815. Two eye-tracking experiments; no self-generation effect found. https://link.springer.com/article/10.1007/s00426-022-01696-2
  7. Cushing, C., and Bodner, G. E. (2022). Reading aloud improves proofreading (but using Sans Forgetica font does not). Journal of Applied Research in Memory and Cognition, 11(3), 427 to 436. Data at OSF. Sample sizes not verifiable from accessible materials. Funded by Flinders University seed and thesis awards. https://psycnet.apa.org/record/2022-07473-001
  8. Pilotti, M., and colleagues (2004, 2009, 2012). Series on familiarity and proofreading, in Journal of General Psychology, Reading and Writing, and Journal of Research in Reading. Changing the typescript restores error salience.
  9. Giurge, L. M., and Bohns, V. K. (2021). “You don’t need to answer right away!” Receivers overestimate how quickly senders expect responses to non-urgent work emails. Organizational Behavior and Human Decision Processes, 167, 114 to 128. Eight preregistered experiments, 4,004 working adults. Funding unverified. https://www.sciencedirect.com/science/article/pii/S0749597821000480