Skip to content

Why does ChatGPT use so many em dashes?

Updated · 8 min read · by the HidenGPT team

Nobody has shown why ChatGPT used so many em dashes, and OpenAI's public statements cover the fix, not the cause. The leading explanations are guesses: the long dash is common in the edited, formal prose models learn from, and training on human ratings may have rewarded answers that looked polished.

The fix is on record. Since GPT-5.1 in November 2025, ChatGPT follows a Custom Instructions line against em dashes much more reliably, according to OpenAI, and a September 2026 study measured GPT-6 Astra using about one-eighth as many as human writers. If a draft still has them, replace each one with a comma, a colon, a period or parentheses.

ChatGPT and the long dash: what is on record and what is a guess (as of 10 October 2026)
ClaimStatusWhere it comes from
Chatbots used the long dash far more than non-professional writers, often where a person would put a comma, colon or parenthesesObservedWikipedia's editors, from AI text added to Wikipedia
Those dashes usually had a space on each side, which most style guides don't ask forObservedWikipedia's editors
Since GPT-5.1, ChatGPT follows a Custom Instructions line against em dashesStated by OpenAI's chief executive in November 2025; Ars Technica's informal test agreedArs Technica report; OpenAI release notes on instruction following
The same request typed in the middle of a chat is less reliableReported by users, not measuredArs Technica report
GPT em dash use fell sharply from GPT-5 to GPT-5.6 Sol and stayed low in GPT-6 Astra, at about one-eighth of the human rateMeasured in one study (web articles, one fixed prompt)Graphite, September 2026
Claude Opus 5 uses them at about the human rate, and Gemini 3.1 Pro has nearly stoppedMeasured in the same studyGraphite, September 2026
Asked for polished writing, a model imitates formal, news and editorial prose, where the dash is commonPlausible, not tested for the dashArs Technica's best-supported explanation
Human raters scored answers with dashes higher because they looked sophisticatedGuessArs Technica calls it speculation
Models picked up the habit from 19th-century books, when dash use in English peakedGuessOnline theory, reported by Ars Technica
Models copied Medium, which turns two hyphens into a long dash automaticallyGuessOnline theory, reported by Ars Technica
The actual cause inside OpenAI's trainingUnknownNot explained in OpenAI's release notes or Help Center as of 10 October 2026

What is on record

Three things are documented.

The pattern was real. Wikipedia's editors, who clean up AI-written text on the encyclopedia, describe chatbots using the long dash far more than non-professional writers in the same genre, often in spots where a person would use a comma, parentheses or a colon, and often to punch up a clause the way sales copy does. They also noticed the dashes usually had spaces around them.

OpenAI changed the behavior. In mid-November 2025, days after GPT-5.1 arrived, chief executive Sam Altman posted on X that ChatGPT now does what it should when your custom instructions tell it not to use em dashes, as Ars Technica reported. OpenAI's release notes for GPT-5.1 say the models follow custom instructions and style preferences more reliably. Neither statement says why the dashes were there in the first place.

The habit has faded in newer GPT models. Graphite, a growth agency, compared 10,000 human web articles with 90,000 articles from nine models on the same topics, published in September 2026. Em dash use fell sharply from GPT-5 to GPT-5.6 Sol and stayed low in GPT-6 Astra, at about one-eighth of the human rate. ChatGPT's paid plans began moving to GPT-6 Sol on 7 October 2026, and no measurement of that model had been published when this page was written.

One small irony: OpenAI's own help page on ChatGPT personalities, read for this page, still shows sample replies with long dashes in them.

What is only guessed

Ars Technica's reporter put it plainly: no one knows precisely why language models overuse the dash. The explanations in circulation rank roughly like this.

Most plausible: models learn from huge amounts of edited writing, such as news, essays and editorials, where the dash is common, and a request for a professional tone pulls them toward that average style. It matches the one thing that is certain about training, that models reproduce patterns common in their data, but nobody has tested it for the dash in particular.

Possible, unproven: during training on human feedback, raters may have preferred answers with dashes because they read as more sophisticated. No rating data has been published to check this.

Weaker: the 19th-century theory, which leans on a 2018 study (cited by Ars) finding that dash use in English peaked around 1860, and the Medium theory, that models copied a blogging site that converts two hyphens into a long dash.

If you read a confident one-line answer somewhere else, it is one of these guesses stated as fact.

How to stop it in ChatGPT

Put the rule in Custom Instructions, not in a single chat. On the web and desktop app, open Settings, choose Personalization, make sure Enable customization is on and type into the Custom Instructions field. On iOS and Android it is Settings, then Customize ChatGPT. OpenAI's help page says the instructions apply to all your chats, and the field holds 1,500 characters on the Free and Go plans and 5,000 on Plus, Pro, Business, Enterprise and Education.

A line that works, 139 characters long: “Never use em dashes or en dashes. Use a comma, colon, period or parentheses instead. Hyphens only inside compound words such as well-known.”

Then test it. Ask for a 300-word piece and search the result for the long dash by copying one into the Find box. As Ars Technica explained, an instruction makes a dash less likely, not impossible, so check anything you plan to send.

The Base style and tone setting won't do this for you. OpenAI's personality page says that for an email draft or a social post, ChatGPT matches tone to your instructions and the context, not necessarily to the personality you picked. If you call the model through the API, put the same line in the system message.

How to fix a draft that already has them

Look at what each dash is doing, then pick the mark that does that job. In the examples, [long dash] stands for the character.

An aside in the middle of a sentence takes commas, or parentheses if it is minor. “The report [long dash] all 40 pages of it [long dash] arrived late” becomes “The report, all 40 pages of it, arrived late.”

An explanation or a list at the end takes a colon. “We changed one thing [long dash] the price” becomes “We changed one thing: the price.”

Two full sentences joined by a dash become two sentences. “Sales doubled [long dash] nobody expected it” becomes “Sales doubled. Nobody expected it.”

A dramatic turn is a different problem. “It isn't a phone [long dash] it's a lifestyle” was built for effect, and a comma keeps the effect. Rewrite it as a plain claim, or cut it.

A range of numbers, which takes the shorter en dash in print, reads fine in words: “pages 10 to 14”, “9 a.m. to 5 p.m.”

A rule of thumb for a chatbot draft: if more than one sentence in ten has a dash, fix them all, then put back the one you would really have written.

Is the long dash still a giveaway in 2026?

Less than it was. Wikipedia's field guide now files the em dash under historical indicators, dated 2022 to September 2026, and says it means something only alongside other signs. Its editors add, without a source so far, that newer models avoid the dash so strictly that a page of nothing but commas and periods has started to look machine-made as well. Graphite's authors reach a similar view from their numbers: GPT and Gemini may have overcorrected.

So a dash proves nothing about who wrote a text, and neither does the lack of one. Writers who have always liked the dash can keep it. What reads as generated is a dash in every paragraph, doing the same dramatic job each time.

As an example of a house rule, HidenGPT tells its drafting model to avoid the long dash almost entirely, and its red pen flags a draft with three or more at a density above one per 120 words, then swaps them for commas, periods or colons when it rewrites. That is a style choice for readability, not a detector rule.

Common mistakes

  • Replacing every long dash with a hyphen: a hyphen joins words, so use a comma, colon or period.
  • Asking once in a chat and assuming it sticks: put the rule in Custom Instructions, then test the next draft.
  • Swapping every dash for a semicolon: a page of semicolons reads just as mechanical, so mix commas, colons and periods.
  • Changing the mark but keeping the dramatic turn: if the sentence was built for a reveal, rewrite it as a plain claim.
  • Treating a dash as proof of AI: plenty of writers use it, and the newest GPT and Gemini models barely do.

Questions

Why does ChatGPT use so many em dashes?

Nobody has shown the cause, and OpenAI has not explained it publicly. The best-supported guess is that models copy the formal, edited prose they learn from, where the long dash is common; a reward for polished-looking answers during training is a second, unproven guess.

Why does ChatGPT use em dashes instead of commas?

Wikipedia's editors observed exactly that: chatbots put long dashes where people would use commas, parentheses or colons, often for punch. Why is unknown, and the practical fix is to tell it which marks to use instead.

How do I permanently remove em dashes from ChatGPT?

Add a line to Custom Instructions (Settings, then Personalization) such as “Never use em dashes; use commas, colons or periods instead.” OpenAI said in November 2025 that GPT-5.1 follows these instructions more reliably, but check each draft, since an instruction lowers the odds rather than ruling a dash out.

Does ChatGPT still use em dashes in 2026?

Far less in recent versions. Graphite's September 2026 study measured GPT-6 Astra at about one-eighth of the human rate in web articles; earlier GPT versions used them much more, and Claude Opus 5 uses them at about the human rate.

Is an em dash a sign of AI writing?

On its own, no. Wikipedia's editors now list it as a historical indicator that only counts alongside other signs, and many careful writers use it.

What is the difference between an em dash, an en dash and a hyphen?

Length and job. The em dash is the longest and marks a break, such as an aside, a sudden turn or an explanation; the en dash is shorter and usually joins ranges like 10 to 14; the hyphen is the shortest and joins words, as in well-known.

Why does ChatGPT keep using em dashes after I tell it to stop?

Because an instruction shifts the odds instead of setting a hard rule, as Ars Technica explained. Requests typed mid-chat were reported to work less well than Custom Instructions, so put the rule there and check each draft.

Sources

  1. Wikipedia: Signs of AI writing (editors' field guide, em dash section) (checked October 10, 2026)
  2. Ars Technica: Sam Altman celebrates ChatGPT finally following em dash formatting rules (14 November 2025) (checked October 10, 2026)
  3. OpenAI Help Center: ChatGPT release notes (checked October 10, 2026)
  4. OpenAI Help Center: ChatGPT Custom Instructions (checked October 10, 2026)
  5. OpenAI Help Center: Customizing Your ChatGPT Personality (checked October 10, 2026)
  6. Graphite research: AI Tells (16 September 2026) (checked October 10, 2026)