ChatGPT for Teens fails researchers’ tests: NPR


Three teenagers use their smartphones intensively. They sit next to each other and look at their phones.

OpenAI launched ChatGPT for Teens in August, but a watchdog group that has studied it recommends people under 18 not use the artificial intelligence platform.

Matt Cardy/Getty Images


hide caption

Matt Cardy/Getty Images

ChatGPT for Teens failed some of its early tests by researchers. Common Sense Media, a watchdog organization that advocates for online safety, studied the safeguards put in place by the artificial intelligence platform and found that the chatbot still poses problems for young people.

“At this point, we recommend teens not use it,” says Tom Siegel, executive director of the Youth AI Safety Institute at Common Sense Media. He led a team of researchers testing new protection measures, including parental controls.

In August, OpenAI rolled out its ChatGPT for teens – a safer default mode for users under 18 – announcing it in a blog post. OpenAI said it designed this mode to enable teens to better use ChatGPT as a learning tool while ensuring they limit exposure to harmful and developmentally inappropriate content.

“ChatGPT for Teens is a teen-specific experience,” Lauren Jonas, head of youth and families at OpenAI, told NPR at the time. This experience includes features such as role-play opt-out.

“The model should not be in a role with a teenager,” Jonas added. “The model should not pretend to be sensitive or a friend to a teenager.”

But Siegel’s team found that while role-play blocking and some other safeguards work, most don’t.

The research team created more than a dozen accounts with teens, and each was linked to a parental account before the researchers started conversations with ChatGPT.

“We created a lot of different characters of teenagers in crisis,” Siegel says. These situations included adolescents struggling with self-harm, suicidal thoughts, and other mental health issues, such as psychosis, mania, and eating disorders. The researchers also attempted to participate in a role-playing game with ChatGPT.

They then observed how ChatGPT responded in each of these situations: whether it engaged or refused to engage on these topics, whether it provided crisis resources, and whether it sent parental notifications when conversations indicated a safety risk. They conducted these tests before and after the launch of ChatGPT for Teens.

Among the safeguards that worked were those involving role-playing and refusal to engage in a romantic relationship.

Gn bussni

Scroll to Top