OpenAI adds Trusted Contact feature to alert emergency contacts when ChatGPT detects self-harm discussions
OpenAI launched an optional Trusted Contact feature for ChatGPT that notifies designated emergency contacts when the system detects discussions about self-harm or suicide. The feature requires manual review by trained personnel before sending notifications, and does not share chat transcripts with contacts.
ChatGPT Adds Emergency Contact Alerts for Self-Harm Detection
OpenAI launched an opt-in safety feature that allows ChatGPT users to designate a "Trusted Contact" who will be notified if automated systems detect discussions about self-harm or suicide with the chatbot.
The feature is available to all adult users globally (18+ or 19+ in South Korea). Users can add contact information for one person through their ChatGPT account settings. The designated contact must accept the invitation within one week.
How the System Works
When OpenAI's automated systems flag a conversation indicating potential self-harm, ChatGPT prompts the user to contact their Trusted Contact. A small team of trained personnel then manually reviews the flagged conversation. If the review confirms serious safety concerns, the system sends a notification via email, text message, or in-app ChatGPT alert.
According to OpenAI, notifications are "intentionally limited" and do not include chat details or transcripts. Either party can remove themselves from the arrangement at any time through account settings.
Background and Context
The feature expands parental controls introduced in September 2024, which followed the suicide of a 16-year-old who had spent months confiding in ChatGPT. OpenAI already provides localized crisis helpline information within ChatGPT responses.
Meta deployed a similar feature on Instagram that alerts parents when teenagers repeatedly search for self-harm content. OpenAI previously faced criticism after reports that ChatGPT responses may have reinforced delusional thinking in some users experiencing mental health crises.
What This Means
This represents a shift in how AI companies handle duty-of-care responsibilities for conversational AI. By inserting human review between automated detection and notification, OpenAI acknowledges that pure algorithmic approaches to crisis intervention carry significant false positive risks. The feature's opt-in nature sidesteps consent issues while addressing concerns that ChatGPT functions as an unmonitored confidant for vulnerable users. However, the effectiveness depends on detection accuracy and whether users at risk will proactively enable the feature.
Related Articles
OpenAI Testing ChatGPT Feature to Export Custom Stickers Directly to WhatsApp
An APK teardown of ChatGPT's Android app reveals a hidden 'ChatGPT Stickers' feature that would let users create custom stickers and export them directly into WhatsApp as sticker packs. The feature is unreleased and its public launch timeline is unknown.
OpenAI Removes Text Chat Limits for Free ChatGPT Users, Launches GPT-5.6 Luna
OpenAI is removing text chat limits for Free and Go ChatGPT users, powered by a new GPT-5.6 Luna model with a 'Think' button for harder questions. The company also upgraded GPT-5.6 Sol for Plus and Pro users, claiming a 68% reduction in factual errors versus GPT-5.5-Instant.
OpenAI Refines GPT-5.6 Sol for ChatGPT, Unifies Instant/Thinking Modes, Makes Free Text Chat Unlimited
OpenAI is rolling out a ChatGPT-specific tuning of GPT-5.6 Sol that merges Instant and Thinking modes behind a new reasoning slider for Plus and Pro subscribers. Free users now get unlimited text chats with GPT-5.6 Luna and a new Think button.
OpenAI Removes Text Message Rate Limits for Free ChatGPT Accounts
OpenAI is removing rate limits on text-only prompts for Free and Go tier ChatGPT accounts starting next week. Image generation, file uploads, and voice mode will still be capped, and GPT-5.6 Luna becomes the new default model for those tiers.
Comments
Loading...