Logo
FrontierNews.ai

OpenAI's New ChatGPT for Teens Has Safety Features, But Experts Say Parents Shouldn't Rely on Them Alone

OpenAI has released a specialized version of ChatGPT designed for teenagers aged 13 to 17, featuring enhanced protections against self-harm, violence, and sexual content, but experts caution that the guardrails are not foolproof and parents should maintain active oversight.

The new ChatGPT for Teens launched on Tuesday as an integrated "mode" within the standard ChatGPT platform. According to OpenAI, the version includes stronger protections around content related to suicide, self-harm, and romantic or sexual conversations, while also providing academic support for homework and studying. The system uses age-detection technology to identify younger users and automatically apply these restrictions.

How Does ChatGPT for Teens Identify and Protect Young Users?

The system operates by analyzing user behavior patterns and responses to estimate age, a process known as age verification or age gauging. If ChatGPT suspects a user has misrepresented their age during signup, it automatically switches them into Teens mode. The platform monitors conversations in real time, looking for warning signs that might indicate a user is heading toward harmful territory.

  • Content Blocking: The system prevents exposure to content involving self-harm, violence, graphic or sexual material, and eating disorders, while also flagging dangerous activities.
  • Real-Time Intervention: When the AI detects concerning conversation patterns, it stops the interaction and redirects users toward mental health resources or emergency services.
  • Academic Features: The version includes tools designed to support learning while making it harder for students to use the tool to cheat on assignments.

"If it thinks that you're a kid, it will offer up a different experience. It will put in more restrictions. It will prevent you from seeing content or being exposed to content that revolves around self-harm or violence or graphic or sexual content," explained Carmi Levy, a technology analyst.

Carmi Levy, Technology Analyst

Can the Safety Features Actually Be Bypassed?

Despite these protections, security researchers have found that the guardrails are not impenetrable. Levy noted that people attempting to bypass the restrictions for research purposes have succeeded in doing so with relatively simple prompting techniques. By changing a few answers or adjusting how they phrase requests, users can circumvent the limitations that OpenAI has built in.

This vulnerability highlights a fundamental challenge in AI safety: the tension between making a tool useful and making it truly secure. Levy was direct about the implications, stating that he would not characterize the system as "100 percent safe." He described this as "the catch-22 of the entire AI industry," acknowledging that no current AI safety measure is completely foolproof.

What's the Real Risk for Parents?

Experts worry that the existence of these safety features may create what Levy calls a "perverse impact," leading parents to assume their children are safer and reduce their own oversight of online activities. This false sense of security could actually increase risk rather than decrease it. Parents may become less vigilant precisely because they believe OpenAI has their children's best interests at heart.

"Parents should not assume that OpenAI has their kids' best interests at heart. They should also not assume that these technologies are perfect. They will miss certain things, and they will also engage in what we call false positives," Levy stated.

Carmi Levy, Technology Analyst

The system can also produce false positives, incorrectly flagging adults as children due to AI errors in age detection. This means legitimate users may experience unnecessary restrictions, while simultaneously, some younger users may slip through the safeguards. Levy emphasized that parents still need to maintain vigilance and not outsource child safety entirely to automated systems.

The launch of ChatGPT for Teens reflects OpenAI's attempt to address growing concerns about AI's impact on young people. However, the technology's limitations underscore a broader principle: AI tools can be helpful supplements to human judgment, but they cannot replace parental involvement and critical thinking about online safety.