Executive Guide
Beyond "I Can't Help With That": How Child Safety Experts Evaluate AI Chatbot Safety
- Author
- Aziz Shuaib Ausi
- Published
- August 11, 2026
- Reading time
- 1 min
- Publication type
- Executive Guide
- Availability
- Open access
Executive Summary
A recent study highlights significant shortcomings in current AI chatbot safety evaluations concerning youth, particularly in high-stakes situations. Existing methodologies often fail to align with real-world harms experienced by young users and rely on unvalidated assumptions about appropriate chatbot responses. This suggests a critical gap in assessing the practical safety of AI systems that youth increasingly use for social and emotional support.
A recent study highlights significant shortcomings in current AI chatbot safety evaluations concerning youth, particularly in high-stakes situations. Existing methodologies often fail to align with real-world harms experienced by young users and rely on unvalidated assumptions about appropriate chatbot responses. This suggests a critical gap in assessing the practical safety of AI systems that youth increasingly use for social and emotional support.
Why it matters
The increasing reliance of youth on AI chatbots necessitates robust safety protocols, yet current evaluation methods are demonstrably insufficient. This creates significant operational and reputational risks for entities deploying or regulating AI systems, as potential harms to vulnerable populations may go undetected, undermining trust and requiring costly retrospective interventions.
Key insights
- Youth are increasingly using AI chatbots for social and emotional support.
- Current AI safety evaluations for children do not adequately incorporate real-world harms experienced by youth.
- Existing evaluations rely on unvalidated assumptions about appropriate AI outputs, such as refusal to engage.
- Evaluations typically focus on detecting adversarial prompts or surface-level harms, overlooking more subtle risks.
- These limitations mean current evaluations may fail to identify AI responses that are harmful to youth in practice.
- The study involved interviews with 19 practitioners, including social workers and therapists, who work directly with vulnerable youth, to understand the limitations of current evaluation practices.
Source
arXiv — Computers and Society — https://arxiv.org/abs/2608.07902
Related publications
Previous
3 AI-Related Questions for Notre Dame’s Sonia Howell
Next
Hardware is an AI Ethics Problem: Expert Visions for a Sustainable and Equitable Semiconductor Industry
Humour as Resistance: Visceralizing the Environmental and Social Impact of AI through Humour-based Creative Practices
Executive Guide
Representational Equality in Cross-country Value Simulation: A Systematic Analysis of Large Language Models
Executive Guide
Data Findability, Governance, and Community Engagement for M\=aori Research Data Sovereignty
Executive Guide
Flow-by-Flow:Content-Judgment Bypass for Governing AI Output in High-Loss Domains
Executive Guide
Automating Freshman Course Placement and Registration: A Case Study
Executive Guide
Rethinking Higher Education: From Fixed Curricula to Learnity Graphs
Executive Guide
Download & citation
Cite this publication (APA 7)
Aziz Shuaib Ausi (2026). Beyond "I Can't Help With That": How Child Safety Experts Evaluate AI Chatbot Safety. Executive Guide. Aziz Shuaib Ausi. https://www.azizshuaib.com/verify/ASA-EXG-2026-00110
Verification
This is an authenticated institutional record.
- Verification ID
- ASA-EXG-2026-00110
- Version
- v1.0 · r0
- Issued
- 8/11/2026
- Publisher
- Aziz Shuaib Ausi
- Licence
- All rights reserved. Reproduction requires written permission.