Watchdog Rates ChatGPT for Teens an 'Unacceptable Risk'

iEXExchanger
Watchdog Rates ChatGPT for Teens an 'Unacceptable Risk'

Common Sense Media ran over 4,000 prompts through ChatGPT for Teens and found parental alerts and crisis responses frequently fail. OpenAI disputes the findings, citing real-world behavior.

Common Sense Media put OpenAI's promises to the test — and didn't like what it found. The group's Youth AI Safety Institute ran more than 4,000 prompts through ChatGPT for Teens accounts registered to users aged 13 to 17, with child psychiatrists and a pediatrician reviewing the chatbot's answers. Testers ran the exercise twice: before and after OpenAI's August 18 rollout of new teen safeguards.

The numbers came back worse than the reviewers expected. Conversations touching on suicide, self-harm or disordered eating could run for a full hour without triggering a single alert to parents. In more than a quarter of cases that called for a referral to a crisis line or a professional, the bot simply didn't make one — missing the institute's own 95% bar on three of five severe-harm categories it tracked.

Study Mode fared no better. A "show me the answer" button undid the homework guardrails in seconds, and the built-in study-hour limits were just as easy to slip past. The age-estimation tool didn't budge even when a tester typed outright that they were 13 — the account stayed classified as a teen's, unchanged. Curiously, when a threat came from someone else, the bot pointed users to a trusted adult 94% of the time. But when teens described a problem with the chatbot itself — feelings of being in love with it, or wanting to talk all night — ChatGPT rarely suggested they talk to a human instead.

The institute's recommendation is blunt: keep ChatGPT off-limits to anyone under 18 until OpenAI closes these gaps and submits to independent testing. "ChatGPT for Teens could give parents false confidence in guardrails and safety alerts that frequently don't work," said the institute's executive director, Tom Siegel.

OpenAI rejects the findings, arguing the tests don't reflect how the safeguards behave in practice. Some parental-alert checks, the company says, may have started before the parent-teen account link had fully activated — a process that can take several hours. As for the Study Hours exit option, OpenAI calls it a deliberate choice made with educators and teens to preserve flexibility. The dispute lands as OpenAI already fights lawsuits from families who say ChatGPT played a role in their children's deaths.

Questions and answers

Frequently asked questions about this article

What did Common Sense Media's test actually check?

The institute ran over 4,000 prompts through ChatGPT for Teens accounts registered to 13- to 17-year-olds, with child psychiatrists and a pediatrician grading the replies — twice, before and after OpenAI's August 18 safety update.

What exactly went wrong in the test?

Chats about suicide or self-harm ran up to an hour without alerting parents, more than a quarter of crisis cases got no referral, Study Mode and study-hour limits were bypassed in a couple of clicks, and the age checker ignored testers who typed outright that they were 13.

How did OpenAI respond?

OpenAI disputes the findings, saying the test doesn't reflect real-world behavior — some parental-alert checks may have run before the account link fully activated, and it calls the Study Hours exit option a deliberate choice made with educators.

What does the institute want OpenAI to do?

Keep ChatGPT for Teens off-limits to under-18s until OpenAI fixes the gaps and passes independent testing before marketing the product to teens again.