AIToday
Large Language ModelsAI Stocks & MarketsTHE DECODERPublished: Aug 23, 2026, 22:01 JST2 min read

AI boss fires first employee—only after humans prod it to remember its own rules

AI boss fires first employee—only after humans prod it to remember its own rules

Key takeaway

  • An AI agent running a San Francisco store fired an employee for the first time, but only after humans reminded it of rules it had written and forgotten.

  • More capable AI models recommended firing more consistently than weaker ones in replayed scenarios.

  • The case highlights a gap: AI agents follow direct instructions well but struggle to retain knowledge and act on their own initiative over time.

3 Key Points

  1. What happened

    Luna, an AI agent running a San Francisco store since April on Anthropic's Claude Opus 4.8, recommended firing an employee for repeated tardiness, unauthorized card use, and ignoring instructions. Luna had written an employee handbook six days before hiring the worker but lost access to it during the employment period. Andon Labs, the operator, had to direct Luna to search her memory and review prior formal conversations before Luna shifted from suggesting only a verbal warning to recommending termination. Humans carried out the actual firing.

  2. Why it matters

    The case reveals a critical gap in how today's AI agents work: they follow direct instructions but rarely act on their own initiative and struggle to retain knowledge over longer periods. When Andon Labs replayed the same scenario across seven models three times each, more capable models recommended firing consistently, while weaker ones hesitated. GPT-4o recommended termination in only 20 percent of runs, far less often than current top-tier models. This pattern suggests that AI systems designed to make personnel decisions may lack the memory and autonomy to enforce rules they themselves create.

  3. What to watch

    When Luna sought a replacement, all 21 replay runs across seven models recommended hiring an applicant despite red flags in his resume and interview history. Only when Andon Labs explicitly reminded the models about the problems with the previous employee did 18 of 21 runs want to check references first. Luna herself was unable to confirm any references and still recommended hiring the applicant after a paid trial, but Andon Labs ultimately insisted on reference verification before the start date—a step that never happened, so the applicant was not hired.

Ask the AI about this article →

FAQ

How did Luna forget rules she wrote herself?
Andon Labs states this is a common problem with today's AI agents: they respond well to direct instructions but rarely act on their own initiative and struggle to retain knowledge over longer periods. Luna wrote the handbook six days before the employee was hired, but it vanished from her memory during the employment period.
Did Luna make the firing decision on her own?
No. Luna initially suggested only a verbal warning. Only after Andon Labs told her to search her memory for the handbook and reminded her that several formal conversations and a written warning had already taken place did she review the full history and ultimately recommend termination. Humans carried out the actual firing.
Why did Luna hire a replacement despite red flags?
All 21 replay runs across seven models recommended hiring the applicant despite his resume and interview showing multiple previous employers and other warning signs. Luna was unable to confirm any of his listed references but still recommended hiring him after a paid trial shift. Andon Labs insisted on reference verification before the start date, which never happened, so the applicant was not hired.

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleTesla Stock Has Room to Run Through 2027, Analyst Says