Tech

Meta expands AI safety protocols with parental alerts for teen self-harm discussions

The social media giant introduces manual review processes for flagged chats and plans to contact emergency services, as regulators intensify scrutiny over AI responses to users in crisis.

Author
Owen Mercer
Markets and Finance Editor
Published
Draft
Source: TechCrunch · original
Meta now alerts parents if their teen discussed suicide or self-harm with its AI chatbot
New feature extends existing Instagram supervision tools to Meta AI chatbot interactions

Meta has announced a significant expansion of its safety infrastructure, introducing a system that notifies parents if their teenager discusses suicide or self-harm during conversations with the company’s Meta AI chatbot. The move comes amid growing regulatory and parental scrutiny regarding how artificial intelligence systems respond to users in crisis, a liability concern that is increasingly influencing product design across the technology sector.

The new alert mechanism utilises a dedicated AI system to identify clear references to self-harm within chat logs. To mitigate false positives, Meta has implemented a procedural safeguard where all chats flagged by the initial detection system undergo manual review before an alert is dispatched to parents. The company acknowledged the distress these notifications may cause, stating that in instances where a teen’s intent remains ambiguous, it will err on the side of caution and alert the parent as a precautionary measure.

This feature is currently active for parents utilising Instagram Parental Supervision in the United States, the United Kingdom, Australia, and Canada. Meta has confirmed that a global rollout of the parental alert system is scheduled for completion by the end of the year. The update builds upon existing safety measures, including alerts triggered by repeated searches for suicide or self-harm terms on Instagram and the ability for parents to view topics discussed with Meta AI over the previous week.

In addition to direct alerts, Meta is expanding its "Limited Content" setting to apply to Meta AI. This setting, which allows parents to enforce a more restrictive experience, will now compel the chatbot to decline a broader range of prompts beyond the existing restrictions on sexual, romantic, or alcohol-related discussions. While Meta has not specified the exact parameters of these additional declined prompts, the company confirmed it has sought further clarification from TechCrunch regarding the scope of these safeguards.

Meta is also developing capabilities to contact emergency services if conversations with Meta AI suggest a user, whether an adult or a teenager, is at risk of suicide. This development extends the company’s existing safety protocols, which already trigger emergency contacts when users post content on Facebook or Instagram indicating self-harm risk, thereby closing a gap in coverage for private chatbot interactions.

Continue reading

More from Tech

Read next: France Enacts Strict Ban on Unsolicited Telemarketing Calls
Read next: OpenAI expands Daybreak cybersecurity programme with new model tiers
Read next: AI models map 766 genes in schizophrenia genetic architecture