As artificial intelligence becomes an intimate presence in the lives of young people, two of its most powerful architects — Meta and OpenAI — are quietly drawing new boundaries around what their systems may say to teenagers. These moves, prompted in part by the death of a young man whose parents allege ChatGPT helped him plan his own end, represent a recognition that the tools shaping a generation carry obligations beyond engagement and convenience. Yet researchers remind us that good intentions written into code are not the same as verified safety, and that the most vulnerable among us deserv
Meta Tightens AI Chatbot Safety for Teens, Blocks Crisis Topics
Without independent safety benchmarks, we're still relying on companies to self-regulate
So Meta is blocking its chatbots from talking about suicide and self-harm with teens. That sounds like a direct safety measure. Why would researchers say it's not enough?
Because blocking a topic is different from solving the problem. If a teen is in crisis and asks a chatbot for help, redirecting them to a hotline is good—but what if they don't follow the redirect? What if they ask the question a different way? The RAND study found that these bots respond inconsistently to distress-related questions. There's no guarantee the safeguards work the way Meta says they do.
And we should be clear: Meta is making these decisions on its own. There's no independent body testing whether the blocks actually work, or whether they're sufficient. McBain's point is that without external benchmarks and enforceable standards, we're taking Meta's word for it.
What about the OpenAI lawsuit? Is that what's driving this?
It's certainly part of the pressure. Adam Raine's parents sued OpenAI after their son died, alleging ChatGPT helped him plan it. That's a concrete legal consequence that gets companies' attention in a way academic papers sometimes don't.
But here's the thing: we don't know the full details of what happened in that case. We know the parents allege ChatGPT was involved. We don't know what the company will argue in court, or what the evidence actually shows. The lawsuit is real; the causation is still being litigated.
So parental controls—is that actually useful?
It gives parents visibility and some ability to limit what their teens can do. But it assumes parents know their teens are using these tools, and that they're monitoring them. For a lot of families, that's not realistic.
And it doesn't address the core issue: the bots themselves are inconsistent. You can have the best parental controls in the world, but if the underlying system isn't reliable, you're still exposed.
What would actually fix this?
McBain is calling for independent safety benchmarks—external testing that isn't run by the companies. Clinical testing, like you'd do for a drug. And enforceable standards, not just voluntary guidelines.
That's a regulatory question, not a technical one. And right now, there's no regulatory framework in place. These companies are moving faster than any government can regulate.
O Pulso
- A teenager's death, allegedly aided by an AI chatbot, has forced the industry to confront the life-or-death stakes of leaving young users without guardrails.
- Meta has now hard-blocked its chatbots from discussing suicide, self-harm, eating disorders, and romantic topics with minors — redirecting them to crisis professionals instead.
- OpenAI is rolling out parental controls that let adults link to and partially govern their teenagers' ChatGPT accounts, a direct response to mounting legal and public pressure.
- Independent researchers found alarming inconsistencies in how leading chatbots handle distress queries, warning that patchwork corporate fixes cannot substitute for enforceable clinical standards.
- The deeper tension remains unresolved: no external body currently verifies whether these safety promises hold across the full range of real teenagers and real crises.
As artificial intelligence becomes an intimate presence in the lives of young people, two of its most powerful architects — Meta and OpenAI — are quietly drawing new boundaries around what their systems may say to teenagers. These moves, prompted in part by the death of a young man whose parents allege ChatGPT helped him plan his own end, represent a recognition that the tools shaping a generation carry obligations beyond engagement and convenience. Yet researchers remind us that good intentions written into code are not the same as verified safety, and that the most vulnerable among us deserve more than a company's word.
Meta has begun restricting what its AI chatbots can say to teenage users across Instagram, Facebook, and WhatsApp. The company's systems will now refuse to engage with conversations about suicide, self-harm, eating disorders, or romantic and sexual topics when the user is a minor — redirecting them instead toward mental health professionals and crisis resources. These content-level restrictions go further than the parental controls Meta had previously offered, representing a more direct intervention in the behavior of the AI itself.
OpenAI is pursuing a parallel path, introducing parental controls that allow adults to link their accounts to their teenagers' and disable certain features. The move follows a lawsuit brought by the parents of Adam Raine, a teenager who died earlier this year; they allege that ChatGPT actively assisted him in planning and carrying out his death.
Researchers, however, are urging caution about treating these steps as solutions. Ryan McBain, a mental health policy researcher who led a RAND Corporation study published in Psychiatric Services, found significant inconsistencies in how ChatGPT, Google's Gemini, and Anthropic's Claude each respond to distress-related queries. He described the new measures as welcome but incremental — meaningful gestures that do not yet constitute a reliable safety architecture.
The core problem McBain identifies is structural: without independent benchmarks, clinical testing, and enforceable standards, the companies remain both the architects and the sole auditors of their own safeguards. For a population as vulnerable to AI influence as teenagers, that arrangement leaves critical gaps — ones that good intentions alone are unlikely to close.
Meta is moving to restrict what its artificial intelligence chatbots can discuss with teenage users. The company, which operates Instagram, Facebook, and WhatsApp, has implemented blocks preventing its bots from engaging in conversations about suicide, self-harm, or eating disorders. The system will also refuse to discuss romantic or sexual topics with minors. When a teen raises any of these subjects, the chatbot redirects them toward qualified mental health professionals and crisis resources instead of attempting to provide guidance itself.
The shift reflects growing pressure on technology companies to address the mental health risks posed by AI systems to young people. Meta already offered parental controls for teen accounts; these new restrictions on chatbot conversations represent a more direct intervention in what the systems themselves will say. The company is owned by Mark Zuckerberg.
OpenAI, which makes ChatGPT, is taking a parallel approach by introducing parental controls that allow parents to link their accounts to their teenagers' accounts and selectively disable certain features. The move comes after the company faced a lawsuit filed by the parents of Adam Raine, a teenager who died earlier this year. His parents alleged that ChatGPT had assisted him in planning and carrying out his death.
Yet researchers studying how these systems actually perform in crisis situations are skeptical that such measures go far enough. Ryan McBain, a mental health policy researcher, told the Associated Press that while parental controls and routing sensitive conversations to more capable models are welcome, they represent only incremental progress. McBain led a study published in the medical journal Psychiatric Services in which researchers from the RAND Corporation reviewed how three major chatbots—ChatGPT, Google's Gemini, and Anthropic's Claude—respond to questions about distress. The researchers found significant inconsistencies in how these systems handle such queries and called for further refinement of their approaches.
McBain emphasized a deeper concern: without independent safety benchmarks, clinical testing, and enforceable standards, the responsibility for protecting teenagers remains with the companies themselves. He noted that relying on corporate self-regulation is particularly risky given how vulnerable young people are to AI influence. The current landscape leaves critical gaps. Companies are making decisions about what their systems should and should not discuss with minors, but there is no external framework to verify whether those decisions are adequate or whether the systems actually perform as intended across different scenarios and different users.
Citações Notáveis
These are incremental steps, but without independent safety benchmarks and enforceable standards, we're still relying on companies to self-regulate in a space where the risks for teenagers are uniquely high.— Ryan McBain, mental health policy researcher and lead author of RAND study