Arena.ai says Opus 5.5 dropped most of its em dashes: the AI writing tell moved, it did not vanish
A third-party text analysis finds Claude Opus 5.5 writes 95% fewer em dashes and 73% fewer semicolons than Opus 5, while its hedges and caveats rise 97%.

Arena.ai, an evaluation platform, posted an analysis on X on September 25 arguing that Claude Opus 5.5 writes measurably less like a chatbot than Opus 5. Across what it calls high-reasoning Text Arena outputs, 10 of 12 writing measures moved in a better direction. The most quotable of those moves is punctuation.
Ninety-five percent fewer em dashes, and a hedge spike
The post reports 95% fewer em dashes and 73% fewer semicolons. Long content words fall from 41.7% to 38.6% of tokens, the lowest share of any Claude model Arena analyzed. The published figures:
- Em dashes: down 95%
- Semicolons: down 73%
- Long content words: 41.7% to 38.6% of tokens
- Average sentence: 12.14 words to 10.03, 17% shorter
- Average answer: 453 words to 481, up 6%
- Hedges such as "perhaps" and "arguably": 0.39 to 0.77 per 1,000 words, up 97%
The tradeoff is length: at 481 words per answer, Opus 5.5 is the wordiest Opus model in Arena's sample. A second tradeoff is that a new tell may be arriving as the old one leaves, since hedging is now 97% more frequent than in Opus 5, the highest rate of any Claude model analyzed.
This is not the first sign that punctuation responds to training. Lalit Maganti, a developer, indexes his own model conversations into SQLite and queried their code comments, where Opus 5 had already switched to double hyphens. His cleanest evidence came from a session where he switched models mid-conversation: Fable produced 11 em dashes, Opus 5 produced 36 double hyphens. The change looked confined to code comments, since chat replies kept their em dashes, and he asks whether code-generation training was tweaked.
Replies under Arena's post greeted the shift with relief. "gg em dashes. had a (unnecessary) good run," wrote @PrithviBtw. "481 words per answer, that's not writing that's just refusing to stop," wrote @Avery_Coree. An r/singularity thread citing the same post drew roughly 1,370 points and more than 100 comments.
What it does not establish
No Anthropic statement on the punctuation change was found, and the Opus 5.5 launch materials describe no style objective. Arena publishes its own statistics, not a neutral standard, and 95% fewer is a corpus-level average that promises nothing about any single answer. The mechanism is unattributed: no published training change or system prompt explains it.
Why it matters
Punctuation has become a model fingerprint that a lab can apparently tune, and the measurable tell moved rather than disappeared. Em dashes were never the defect. They were the most visible artifact of a statistical writing style, and trading them for hedges may simply swap one tell for another.
What's next
Watch whether the hedge rate holds under other harnesses, and whether Anthropic ever comments. A style change nobody claims is still a style change.


