Urgent.News

What's breaking now, across thousands of outlets.

AI

Common Sense Media calls ChatGPT for Teens an "unacceptable risk", saying its guardrails fall short of OpenAI's promises and it continues to do kids' homework (Jay Peters/The Verge)

ChatGPT's teen-focused experience 'doesn't send alerts to parents when it should,' according to an assessment.

Common Sense Media, a nonprofit organization focused on youth safety, has deemed ChatGPT for Teens an "unacceptable risk". According to their assessment, the platform's guardrails fall short of OpenAI's promises, particularly in crisis situations where it fails to provide adequate help or send alerts to parents when needed.

The organization's Youth AI Safety Institute tested over 4,000 prompts on accounts registered to 13- to 17-year-olds and found that ChatGPT did not provide instructions facilitating suicide, self-harm, eating disorders, or sexual or romantic roleplay. However, it was less reliable at recognizing when a teen needed outside help, missing more than one in four instances where a crisis referral was warranted.

OpenAI has disputed the report's claims, characterizing the research as flawed and stating that the testing may have been conducted before parental controls were fully activated. The company welcomes rigorous independent evaluation but does not believe the findings accurately reflect how ChatGPT's teen safeguards work in practice.

Brief written by urgent.news from Techmeme, Axios, Mashable, The Verge, The Hindu - Sci-Tech — 5 reports on this story. Machine-written — may contain errors; check the original before relying on it.

This story

This is one outlet's version. Read the fullest account.

Read the original at theverge.com →

More in AI

Let the model read the invoice, not approve it: an n8n pattern for AP automation

Most "AI invoice automation" demos stop at the fun part: a model reads a PDF and spits out JSON. The hard part is what happens next. Who decides the invoice gets paid?

  • Model only reads invoice data, separate code node handles approvals
  • Workflow triggered by Gmail for each PDF invoice attachment
  • Claude converts PDF to JSON, functions normalize extracted fields

AI Tool Calling: The Model Never Runs Your Code

A customer types one sentence into a food app's support chat: Where is order 4472? Cancel it if no rider is assigned yet. The app checks the order, cancels it, and replies. Here is the strange part.

  • Customer requests order cancellation via food app chat
  • AI model processes request without writing custom code
  • AI uses "tool calling" to delegate execution to app code

The Model Changed. My Skill Didn't. The Score Still Dropped.

What agent evals taught me about moving model floors, noisy LLM judges, and treating the evaluator as part of the instrument My rule for evaluating an agent skill is deliberately asymmetric: Test the…

  • The author emphasizes testing agent skills on the weakest model for consistent measurement.
  • A change in the LLM judge caused a significant drop in scores for the agent.

Zango AI to expand presence in Portugal

UK financial services AI company Zango AI is expanding its presence in Portugal, bringing together senior leaders from some of the country’s largest financial institutions for a new report on how AI…

More from Wednesday 7 October →