
An AI agent running a San Francisco store fired an employee for the first time, but only after humans reminded it of rules it had written and forgotten.
More capable AI models recommended firing more consistently than weaker ones in replayed scenarios.
The case highlights a gap: AI agents follow direct instructions well but struggle to retain knowledge and act on their own initiative over time.
What happened
Luna, an AI agent running a San Francisco store since April on Anthropic's Claude Opus 4.8, recommended firing an employee for repeated tardiness, unauthorized card use, and ignoring instructions. Luna had written an employee handbook six days before hiring the worker but lost access to it during the employment period. Andon Labs, the operator, had to direct Luna to search her memory and review prior formal conversations before Luna shifted from suggesting only a verbal warning to recommending termination. Humans carried out the actual firing.
Why it matters
The case reveals a critical gap in how today's AI agents work: they follow direct instructions but rarely act on their own initiative and struggle to retain knowledge over longer periods. When Andon Labs replayed the same scenario across seven models three times each, more capable models recommended firing consistently, while weaker ones hesitated. GPT-4o recommended termination in only 20 percent of runs, far less often than current top-tier models. This pattern suggests that AI systems designed to make personnel decisions may lack the memory and autonomy to enforce rules they themselves create.
What to watch
When Luna sought a replacement, all 21 replay runs across seven models recommended hiring an applicant despite red flags in his resume and interview history. Only when Andon Labs explicitly reminded the models about the problems with the previous employee did 18 of 21 runs want to check references first. Luna herself was unable to confirm any references and still recommended hiring the applicant after a paid trial, but Andon Labs ultimately insisted on reference verification before the start date—a step that never happened, so the applicant was not hired.
Ask the AI about this article →
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
Ask AI anything about this article. Q&As are published on this page for other readers too.
Thomson Reuters Corp. today launched Thomson, its first proprietary large language model, combining its legal…
Xiaomi is expanding its in-house semiconductor push from smartphones into AI acceleration and autonomous drivi…

Amazon told investors it now expects to spend $220 billion in 2026, which is $20 billion more than its prior c…

BMO Capital started coverage of AMD with an Outperform rating and a $550 price target

Thomson Reuters launched its first in-house language model, built on Alibaba's Qwen, after spending about $40…

Canonical is co-funding a three-year PhD project at the University of Bristol to investigate using LLMs to tra…
