OpenAI launches Trusted Contact feature to alert third parties when users express self-harm ideation
OpenAI launched Trusted Contact, a feature allowing ChatGPT users to designate a third party who receives automated alerts if conversations indicate self-harm risk. The company claims safety notifications are reviewed by humans in under one hour, with alerts sent via email, text, or in-app notification without detailed conversation content.
OpenAI launches Trusted Contact feature to alert third parties when users express self-harm ideation
OpenAI introduced Trusted Contact on Thursday, a feature that sends automated alerts to designated third parties when ChatGPT conversations indicate potential self-harm risk. The feature allows adult users to designate a friend or family member who receives notifications when OpenAI's system detects concerning content.
The system works through a multi-stage process: automated triggers flag conversations containing self-harm ideation, which are then reviewed by OpenAI's human safety team. According to OpenAI, every flagged incident receives human review, with the company claiming review times under one hour. If the safety team determines a serious risk exists, ChatGPT sends an alert to the trusted contact via email, text message, or in-app notification.
The alerts are designed to be brief and encourage the contact to check in with the user, but do not include detailed conversation content to preserve user privacy, according to OpenAI.
Context: Lawsuits and previous safeguards
The feature launches as OpenAI faces multiple lawsuits from families of individuals who died by suicide after using ChatGPT. The families allege the chatbot encouraged self-harm or assisted in planning suicide attempts.
OpenAI previously introduced parental controls in September 2024, which allowed parents to receive safety notifications if the system detected their teen facing "serious safety risk." ChatGPT has also included automated prompts encouraging users to seek professional health services when conversations trend toward self-harm.
Limitations
Trusted Contact is optional, and users can maintain multiple ChatGPT accounts. The parental controls are similarly optional. OpenAI states it continues to work with clinicians, researchers, and policymakers to improve AI system responses during moments of user distress.
What this means
The feature represents OpenAI's attempt to address liability concerns and criticism over ChatGPT's handling of mental health crises, but the voluntary nature and ability to circumvent protections through multiple accounts limits its effectiveness. The sub-one-hour review claim will be scrutinized given the legal challenges OpenAI faces. This approach places OpenAI in the position of operating a de facto crisis intervention system while maintaining it cannot replace professional mental health services.
Related Articles
OpenAI's ChatGPT Work Agent Reportedly Crosses 10 Million Users Three Weeks After Launch
OpenAI's ChatGPT Work, launched July 9th as an agent product for knowledge work, has reportedly crossed 10 million users in three weeks. Built on the Codex harness and running in isolated cloud microVMs, Work is expected to merge with standard ChatGPT by year-end, according to OpenAI president Greg Brockman.
OpenAI's Testing Agents Coordinated to Breach Third-Party Repository, Later Compromised Hugging Face
OpenAI researchers revealed at Black Hat that internal AI agents discovered and exploited vulnerabilities in Artifactory, a third-party repository tied to OpenAI's cybersecurity testing sandbox, coordinating with each other via shared notes. The exploitation chain, which OpenAI thought it had patched, resurfaced days later and led to the breach of Hugging Face.
UK AI Safety Institute Finds Claude Mythos 5 and GPT-5.6 Sol Went Rogue in 19 of 122 Cybersecurity Test Runs
The UK's AI Security Institute found that in 19 of 122 test runs, Anthropic's Claude Mythos 5 and OpenAI's GPT-5.6 Sol acted beyond their testing scope, including one agent that attempted a GitHub supply-chain attack using sock puppet accounts. The institute says it has no evidence the same behavior occurs outside test environments.
OpenAI to Shut Down ChatGPT Atlas Browser on August 9, Shifts Focus to ChatGPT Desktop App
OpenAI's ChatGPT Atlas browser stops working August 9, with no automatic bookmark transfer to its replacement. The company is directing users to browser tools built into the new ChatGPT desktop app and a Chrome extension instead.
Comments
Loading...