Claude Opus 5.5 Moves Closer to Human Writing but Retains 2,548 Tells
Graphite found Claude Opus 5.5’s writing distribution closer to its human sample than its predecessor’s, even as thousands of recurring words, phrases and patterns remained detectable.

Key takeaways · 3
- 01
Evaluate model writing with broad language distributions rather than relying on a few conspicuous stylistic habits.
- 02
Claude Opus 5.5’s reduced em-dash usage does not eliminate thousands of other detectable patterns.
- 03
Writing behavior can diverge between model generations rather than improving consistently across every provider.
What Graphite Measured
Graphite compared articles from ten models and human writers across 9,974 matched topics. [1] Claude Opus 5.5 recorded a word-distribution divergence score of 0.052 against the human sample, down from 0.064 for Opus 5, a 19% reduction. [1] Graphite found 2,548 words, phrases and sentence patterns that appeared at least twice as often in Opus 5.5 articles as in comparable human articles. [1]
The model produced 0.015 em-dashes per 1,000 words, versus 2.92 for Opus 5, a decline of about 99%. [1] OpenAI’s GPT-6 Astra scored 0.109 on word-distribution divergence, up from 0.101 for GPT-5.6 Sol. [1]
What it means
Claude Opus 5.5’s lower divergence score indicates closer alignment with Graphite’s human sample, but the remaining 2,548 tells show why a single stylistic cue is an inadequate test. The near-disappearance of em-dashes removes one conspicuous habit without making the model’s broader language patterns indistinguishable from human writing.
The comparison with GPT-6 Astra also shows that newer generations do not necessarily move in the same direction on this measure: Opus moved closer to the human sample while Astra’s divergence increased. What the sources don't address: whether these measured differences remain consistent across languages, formats and real enterprise writing workflows.
The results show that removing a recognizable AI habit does not necessarily eliminate broader statistical differences from human writing. Practitioners evaluating generated content should examine multiple linguistic patterns rather than treating one punctuation mark or stock phrase as decisive.
Why it matters
Put this to work — one session a day, built for your industry.
Create a free account for a daily session — eight questions and one real-work challenge, on the news that affects your role.
Start freeHow this developed
1 October 2026
Claude Opus 5.5 Moves Closer to Human Writing but Retains 2,548 Tells
1 October 2026
Event created from source cluster.