Intelligence

ai

Beyond "I Can't Help With That": How Child Safety Experts Evaluate AI Chatbot Safety

Source
arXiv — Computers and Society
Published
Last verified
11 Aug 2026
Confidence
High
Evidence
Original document retained
Reading time
1 min
Country
International
Relevant to
Research & Evidence, Technology & Data, Risk & Compliance

Executive summary

What happened, and why should leadership care?

A recent study highlights significant shortcomings in current AI chatbot safety evaluations concerning youth, particularly in high-stakes situations. Existing methodologies often fail to align with real-world harms experienced by young users and rely on unvalidated assumptions about appropriate chatbot responses. This suggests a critical gap in assessing the practical safety of AI systems that youth increasingly use for social and emotional support.

Why this matters

Why is this strategically important?

The increasing reliance of youth on AI chatbots necessitates robust safety protocols, yet current evaluation methods are demonstrably insufficient. This creates significant operational and reputational risks for entities deploying or regulating AI systems, as potential harms to vulnerable populations may go undetected, undermining trust and requiring costly retrospective interventions.

Key insights

What should be noted from the evidence?

  • Youth are increasingly using AI chatbots for social and emotional support.
  • Current AI safety evaluations for children do not adequately incorporate real-world harms experienced by youth.
  • Existing evaluations rely on unvalidated assumptions about appropriate AI outputs, such as refusal to engage.
  • Evaluations typically focus on detecting adversarial prompts or surface-level harms, overlooking more subtle risks.
  • These limitations mean current evaluations may fail to identify AI responses that are harmful to youth in practice.

Evidence and confidence

How far can this assessment be trusted?

High confidence. Named institution, original document retained and analysis corroborated.

Analysis is prepared editorially by Aziz Shuaib Ausi. The original publication remains the authoritative record, and executive judgement remains entirely human.

Source

Where does this originate?

Reported by arXiv — Computers and Society · International. This briefing summarises the publication for executive use; the document itself is not reproduced here.

Read the original publication