AIToday
Large Language ModelsTechCrunch AIPublished: Oct 2, 2026, 04:00 JST

Opus 5.5 says 'this matters' 116x more than humans

Opus 5.5 says 'this matters' 116x more than humans

3 Key Points

  1. What happened

    Graphite studied frontier models against 10,000 pre-ChatGPT articles and found 13,000 phrases at least twice as common in AI content, with Opus 5.5 using "this matters" 116 times more often than humans.

  2. Why it matters

    The 13,000-phrase count suggests AI writing habits remain detectable at scale even after the best-known tells, such as em-dashes, are removed, and each model version carries its own set.

  3. What to watch

    Whether the overall number of tells keeps holding steady hinges on whether labs can control behaviors in models with billions of parameters; watch whether Anthropic's claim that Opus 5.5 "communicates more naturally than prior models" holds up in future samples.

WHO IT HITSEditors, publishers, and communications teams screening submitted or published prose for AI authorship are the most directly affected, since the study's phrase-level tells give them concrete markers to test. Recruiters and admissions reviewers reading large volumes of personal statements may also find the findings useful for triage.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The study's design matters for how much weight the findings can carry. Graphite used a corpus of 10,000 articles published before ChatGPT was released as a human control, then had each AI model rewrite those articles from summaries to reduce source bias. That setup lets the researchers compare the same underlying content written by humans and by each model, rather than comparing unrelated texts. The 13,000-phrase figure is the study's threshold for a tell — a phrase at least twice as common in AI content as in human content — and it is unusually broad compared with earlier, more anecdotal lists of AI giveaways.

The findings also show that tells shift rather than shrink. All the frontier labs appear to have responded to the criticism that models overuse em-dashes: Opus 5.5 used them 99% less than Opus 5, Astra 88% less than human samples, and Gemini 3.1 Pro has almost eliminated them. Yet Graphite's chief AI officer, Greg Druck, says the overall number of tells is mostly holding steady, with each model version displaying its own quirks. That pattern is notable because Anthropic and OpenAI both marketed their newest releases on more natural writing — Anthropic said Opus 5.5 "communicates more naturally than prior models," and OpenAI said users of the GPT-6 versions of Sol and Luna could expect "more clarity, less jargon, [and] fewer odd turns of phrase." Druck is skeptical that labs can fully control these habits, noting that giant models with billions of parameters have a finite number of tests they can run.

The practical stakes sit with anyone who has to judge whether a piece of prose was written by a person or a model — editors, publishers, recruiters, and admissions reviewers among them. The study suggests the old shortcuts, like scanning for em-dashes or the word "delve," are no longer reliable, since labs have largely stamped those out. What remains to be seen is whether the underlying rate of tells stays steady as models get larger, or whether more targeted fixes can bring individual quirks down without new ones appearing.

FAQ
What words or phrases give away AI writing, according to the study?
Claude Opus 5.5 uses "dependable" 23 times more often than human samples and "this matters" 116 times more often. OpenAI's Astra favors "corrective framing" like "not simply X" or "rather than relying on X," which was over 100 times more common than in human writing.
Did the models stop using em-dashes?
Yes, largely. Opus 5.5 used em-dashes 99% less often than Opus 5, Astra used them 88% less than human samples, and Gemini 3.1 Pro has almost completely eliminated them.
How did Graphite test for AI writing tells?
Researchers started with 10,000 articles published before ChatGPT's release as a human control group, then had AI models rewrite those articles from summaries to compare word and phrase frequencies.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleAmazon Payments bandit lifts funnel conversion high single digits