AIToday
Large Language ModelsAI Business & IndustryOpenAI BlogPublished: Apr 29, 2026, 10:01 JST1 min read

OpenAI strengthens ChatGPT safeguards to detect subtle risks of violence and harm across conversations

3 Key Points

  1. OpenAI has expanded ChatGPT's safety systems to better recognize warning signs of potential violence across long, high-stakes conversations. The work builds on years of model training, evaluations, red teaming, and expert input.

  2. When ChatGPT detects that a user may be planning or carrying out violence, OpenAI revokes access to its services, including disabling the account and banning other accounts of the same user. The company maintains a zero-tolerance policy for using its tools to assist in committing violence. When conversations indicate imminent and credible risk of harm to others, OpenAI notifies law enforcement.

  3. OpenAI uses automated detection systems (including classifiers, reasoning models, hash-matching technologies, and blocklists) to flag potentially concerning activity, then trained human reviewers assess flagged accounts in context. For cases with indicators of serious real-world harm, OpenAI escalates to in-depth investigation using structured criteria.

  4. OpenAI has introduced Parental Controls allowing parents to customize age-appropriate settings for their teen's ChatGPT use without accessing conversations, with automatic notification to parents in rare cases where the system detects acute distress. A trusted contact feature for adult users will soon be introduced, allowing users to designate someone to receive notifications when additional support may be needed.

Ask the AI about this article →

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • DataAgent launches with $10M to auto-fix Kubernetes faultsSiliconANGLE AI · 50m ago
  • SK Hynix custom HBM boosts inference up to 5.15xDIGITIMES Asia · 50m ago
  • Nvidia Earnings: Boring by Design, Avoiding a Consolidated WorldStratechery (Ben Thompson) · 50m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleMeta released Muse Spark, its first large language model from the Superintelligence Lab, as part of a $135 billion capital expenditure plan for the year, up from $72 billion a year ago.