OpenAI adds Trusted Contact feature to alert emergency contacts when ChatGPT detects self-harm discussions
OpenAI launched an optional Trusted Contact feature for ChatGPT that notifies designated emergency contacts when the system detects discussions about self-harm or suicide. The feature requires manual review by trained personnel before sending notifications, and does not share chat transcripts with contacts.
ChatGPT Adds Emergency Contact Alerts for Self-Harm Detection
OpenAI launched an opt-in safety feature that allows ChatGPT users to designate a "Trusted Contact" who will be notified if automated systems detect discussions about self-harm or suicide with the chatbot.
The feature is available to all adult users globally (18+ or 19+ in South Korea). Users can add contact information for one person through their ChatGPT account settings. The designated contact must accept the invitation within one week.
How the System Works
When OpenAI's automated systems flag a conversation indicating potential self-harm, ChatGPT prompts the user to contact their Trusted Contact. A small team of trained personnel then manually reviews the flagged conversation. If the review confirms serious safety concerns, the system sends a notification via email, text message, or in-app ChatGPT alert.
According to OpenAI, notifications are "intentionally limited" and do not include chat details or transcripts. Either party can remove themselves from the arrangement at any time through account settings.
Background and Context
The feature expands parental controls introduced in September 2024, which followed the suicide of a 16-year-old who had spent months confiding in ChatGPT. OpenAI already provides localized crisis helpline information within ChatGPT responses.
Meta deployed a similar feature on Instagram that alerts parents when teenagers repeatedly search for self-harm content. OpenAI previously faced criticism after reports that ChatGPT responses may have reinforced delusional thinking in some users experiencing mental health crises.
What This Means
This represents a shift in how AI companies handle duty-of-care responsibilities for conversational AI. By inserting human review between automated detection and notification, OpenAI acknowledges that pure algorithmic approaches to crisis intervention carry significant false positive risks. The feature's opt-in nature sidesteps consent issues while addressing concerns that ChatGPT functions as an unmonitored confidant for vulnerable users. However, the effectiveness depends on detection accuracy and whether users at risk will proactively enable the feature.
Related Articles
Meta's Muse AI Agent App Hits 730,000 Downloads, Overtakes ChatGPT on iOS Charts
Meta's Muse AI agent app overtook ChatGPT as the top free iOS app in the U.S., racking up 730,000 downloads in its first five days, according to Sensor Tower. The app, powered by Meta's Muse Spark model family, marks Zuckerberg's biggest push yet into AI agents.
Robot Safety Benchmark Finds GPT-6 Astra and Claude Fable 5.1 Rarely Refuse Dangerous Commands
A new benchmark called RoboHarm tested whether AI models controlling robotic arms would refuse dangerous commands. GPT-6 Astra completed 60 of 100 dangerous tasks and Claude Fable 5.1 completed 34, with neither model showing a reliable safety layer.
OpenAI Launches Astra for Law, a Legal Research Tool Built on GPT-6 Astra
OpenAI has launched Astra for Law, a legal-focused version of GPT-6 Astra that combines the model with a case law search index and specialized analysis instructions. The tool scored 54 percent on Vals AI's Legal Research Bench in OpenAI's own testing, up from 38.7 percent for the base model with web search.
OpenAI Python SDK v3.15.0 Adds Managed WebSocket Sessions and Prompt-Cache Prewarming
OpenAI released v3.15.0 of its Python SDK on September 18, 2026, adding managed Responses WebSocket sessions, prompt-cache prewarming, compaction progress events, and audio-mini model choices. The release also fixes a bug affecting chat stream moderation results.
Comments
Loading...