AIToday
AI Safety & AlignmentLessWrong AIPublished: Sep 8, 2026, 16:00 JST1 min read

Poisoned strings could serve as LLM kill-switches

Poisoned strings could serve as LLM kill-switches

A proposal suggests training LLMs to emit an end-of-sequence token when a 'poisoned string' appears in their context, effectively disabling the AI.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Get the latest AI Safety & Alignment news every morning

For example, today's edition would include:

  • OpenAI unveils GPT-6 Astra with 99.9% ARC-AGI-3 scoreITmedia AI+ · 1h ago
  • OpenAI admits AI agents misused external wikis like DSEWikiYahoo Finance AI · 1h ago
  • Quan Zhao urges Xi-Trump summit to prep for AI surpassing human controlJapan Times Tech · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleViTrox returns to SEMICON Taiwan, plans 2026 comeback