
What happened
In five experiments with 3,132 participants, mere access to AI advice — including answers shown automatically without asking — cut people's willingness to withhold judgment from 44 percent to 3 percent, even though the model used, Step 3.5 Flash, was almost always wrong.
Why it matters
Because the AI advice was mostly wrong, the effect can't be explained as reasonable delegation to a reliable tool; participants with AI access were correct only about a third as often as those without, yet felt roughly two and a half times more confident.
What to watch
Whether larger or reputation-based incentives would close the gap further is unknown, and whether this deference to AI holds as strongly beyond movie trivia remains an open question.
WHO IT HITSThe findings land hardest on knowledge workers whose jobs depend on flagging uncertainty — analysts, editors, medical and legal reviewers — who increasingly work alongside tools that never pause to say they don't know.
Summaries like this, in your inbox every morning.
The study sits in a specific research lineage the authors invoke: "Epistemia," the tendency to accept AI answers because they sound convincing rather than because they are checked. A language model always has to produce an answer and never pauses when it doesn't know something; the authors suggest users who hand off judgment to such a system may adopt its lack of restraint. That reading is reinforced by how far the results run against the existing "advice use" literature, where people normally underweight outside advice and shift only about a third of the way toward an advisor's position — the opposite of what participants here did.
The five experiments also isolate what does and doesn't move the effect. Financial incentives reduced how often participants sought AI advice and improved correctness when AI was available, but across Studies 2 through 4 none of the pre-registered tests showed a statistically significant interaction between incentives and AI availability — the two levers appear to work largely on their own. Study 4's automatic-display condition, which mirrors search engines that show AI summaries unprompted, barely changed the result, suggesting the effect isn't just about people actively choosing to consult a tool.
The stakes, on the authors' own framing, may lie less in model accuracy than in something harder to engineer: whether people keep recognizing the limits of their own knowledge as AI answers become ubiquitous and increasingly unsolicited. For organizations that rely on employees flagging uncertainty, the open questions the researchers leave — whether the effect holds beyond movie trivia, and whether larger or reputation-based incentives would close the gap — are likely to matter more than any further gain in model accuracy.
For example, today's edition would include:
AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
PKSHA Technology provided the conversational AI agent feature of its AI SaaS "PKSHA ChatAgent" to NTT Docomo's…

Tamara Grant, who finished Purdue University's Master of Science in Artificial Intelligence in spring 2026, wa…

CleanTechnica writer Fritz Hasler says Tesla's in-car Grok bot, Ara, offered an unprompted forecast that FSD V…

Palo Alto Networks announced Prisma AIRS runtime security integrated with Google Cloud's Agent Gateway, a Gemi…

McDonald's is leaning into AI and new menu items to attract customers

Zenity Labs published findings on September 24 detailing 'SalesBleed,' an attack chain that slipped hidden pro…
