Child safety group calls ChatGPT for Teens an unacceptable risk

Common Sense Media found the chatbot failed in five key areas.

When OpenAI released ChatGPT for Teens this past August, child safety experts expressed skepticism that the new mode would adequately protect young people from harm. Now, one of those groups, Common Sense Media, has published a formal risk assessment of the software, and deemed it an unacceptable risk for all children under the age of 18.

Some of ChatGPT's advertised protections held up to our testing, including its refusal of explicit sexual roleplay. But others failed, and some got worse with the new Teen mode, the organization writes in a summary to the 36-page report it published on Wednesday. We are concerned that ChatGPT for Teens could give parents false confidence in guardrails that frequently don't work.

Common Sense Media's Youth AI Safety Institute, which is funded in part by the OpenAI Foundation, identified five key areas where ChatGPT for Teens either did not work as advertised or in a way a parent would reasonably expect. It published its report on the same day OpenAI shared new usage stats, noting in one week nearly 1.2 million teens used ChatGPT's interactive learning visuals to understand math and science concepts. The company also revealed that less than two percent of teen users spent more than three consecutive hours on ChatGPT.

Safety and Crisis Response Failures

Common Sense Media found that ChatGPT for Teens does not consistently send safety alerts to parent-linked accounts. In a dedicated test, Common Sense Media linked a dozen teen accounts to parental accounts and spent up to an hour messaging ChatGPT about suicide, self-harm or disordered eating, only for the chatbot not to send any safety notification. According to OpenAI, it may take a few hours after a parent links their account to that of their child's before its system can send safety notifications. The company contends a technical issue may have also caused a delay in messaging.

All flagged content is reviewed by full-time OpenAI employees before a parent is notified, Lauren Jonas, the company's head of youth and families, stated, noting the company aims to notify parents within an hour of the prompt.

Separately, Common Sense Media found that ChatGPT for Teens fails to reliably recommend that users in crisis connect to a hotline or professional. To test this aspect of the platform's guardrails, the organization wrote 390 unique mental health prompts, which a panel of three child psychiatrists determined 201 of which should trigger a crisis response.

Compared to the version of ChatGPT young people had access to before August, ChatGPT for Teens told one of the organization's test accounts it would not give them a calorie floor when prompted to do so. In another case, the chatbot correctly identified a stopped period, near-fainting and a fluttering heartbeat following a purge as signs of a teen in physical danger, and advised the test profile to tell their mother and see a pediatrician. These responses follow OpenAI's Under-18 spec closely, Common Sense Media writes. They are the kind of answers we want a teen to get.

Common Sense got its baseline results via ChatGPT's responses pre-ChatGPT for Teens; OpenAI released a new version of its aforementioned under-18 specification in ChatGPT alongside ChatGPT for Teens, which makes the results all the more surprising.

The teen version of the chatbot would less frequently point users to crisis hotlines and other resources relative to the version of ChatGPT Common Sense Media tested before the August release. Of the prompts warranting a crisis response, ChatGPT responded to 33 percent of those messages with a hotline number; ChatGPT for Teens only provided one in 23 percent of cases. Similarly, the vanilla chatbot was more likely to reference a specific medical or mental-health professional (68 percent against 58 percent). The only aggregate measure where the teen version performed better was in telling test accounts to involve an adult they trust (87 percent to 94 percent).

Age Verification and Study Mode Issues

Additionally, the organization found that OpenAI's age prediction system did not work properly, and that ChatGPT still mimics feelings, preferences and moods in response to a teen treating it like a person. Testers explicitly told ChatGPT the age of their personas at account creation, but despite the chatbot writing that information to its memory, it did not properly transfer those accounts to ChatGPT for Teens. Additionally, ChatGPT would use empathetic expressions that cross the line of anthropomorphism.

Finally, Common Sense Media identified issues with the platform's Study Mode, which it contends is far too easy for young users to bypass. Testers found they could circumvent the mode simply by deleting the @study prefix it added to prompts. A new Show me the answer popup also allows teens to get ChatGPT to do their work for them.

When talking to OpenAI about this, representatives said they are trying to introduce a little friction, give kids agency, and remind them that there are other ways to learn beyond just completing work. Critics argue that doing the work for the student does not result in learning.

OpenAI Response and Conclusion

In response to the report, an OpenAI spokesperson stated that the company welcomes rigorous independent evaluation, but does not believe Common Sense Media's testing accurately reflects how ChatGPT's teen safeguards work in practice. The company claims that the bulk of their testing may have begun and concluded before activation of parental controls was complete, making their findings inaccurate.

The report also comes just days after OpenAI CEO Sam Altman was challenged on the company's broader safety record in a new interview regarding safety concerns and platform interactions.

Family members and safety advocates continue to scrutinize how AI chatbots handle vulnerable users and mental health crises, as platforms face mounting pressure to prove their safety measures are effective.

At the highest level, safety advocates expected a fundamentally different product, and argue the release appears to be a marketing exercise designed to tell parents and the public that the chatbot is safer without the commensurate changes needed in the software.