arXiv:2505.10588cs.CYcs.AI2025-05中稿 · ACM FAccT 2025被引 11

评估大模型对00后数字语言的识别能力,发现现有安全系统严重失效。

Understanding Gen Alpha Digital Language: Evaluation of LLM Safety Systems for Content Moderation

  • 用100条00后网络表达构建首个专属数据集
  • 四款主流大模型对隐性欺凌识别率普遍不足50%
  • 引入00后共同研究,揭示代际沟通鸿沟

本研究首次评估AI系统对2010-2024年出生的00后群体(Gen Alpha)数字语言的理解能力。作为首个在AI环境中成长的一代,00后因沉浸式数字互动和语言演变,面临新型网络风险,其受游戏、梗图和AI驱动趋势影响的语言常隐藏有害内容,导致人工与自动审核系统均难以识别。研究基于100条来自游戏平台、社交媒体和视频内容的真实表达,评估GPT-4、Claude、Gemini和Llama 3四款主流模型的检测能力。结果揭示关键理解偏差,直接威胁在线安全。研究贡献包括:(1)首个涵盖00后语言表达的数据集;(2)面向青少年保护的AI审核优化框架;(3)包含AI系统、人工审核员及家长的多视角评估,含00后共研者反馈;(4)分析语言差异如何加剧青少年脆弱性。研究强调亟需重构适配青少年语境的安全系统,尤其当00后因成人不理解其数字世界而拒绝求助时。研究结合00后研究员洞见与系统性学术分析,应对关键数字安全挑战。

原文摘要 · Abstract (English)

This research offers a unique evaluation of how AI systems interpret the digital language of Generation Alpha (Gen Alpha, born 2010-2024). As the first cohort raised alongside AI, Gen Alpha faces new forms of online risk due to immersive digital engagement and a growing mismatch between their evolving communication and existing safety tools. Their distinct language, shaped by gaming, memes, and AI-driven trends, often conceals harmful interactions from both human moderators and automated systems. We assess four leading AI models (GPT-4, Claude, Gemini, and Llama 3) on their ability to detect masked harassment and manipulation within Gen Alpha discourse. Using a dataset of 100 recent expressions from gaming platforms, social media, and video content, the study reveals critical comprehension failures with direct implications for online safety. This work contributes: (1) a first-of-its-kind dataset capturing Gen Alpha expressions; (2) a framework to improve AI moderation systems for youth protection; (3) a multi-perspective evaluation including AI systems, human moderators, and parents, with direct input from Gen Alpha co-researchers; and (4) an analysis of how linguistic divergence increases youth vulnerability. Findings highlight the urgent need to redesign safety systems attuned to youth communication, especially given Gen Alpha reluctance to seek help when adults fail to understand their digital world. This study combines the insight of a Gen Alpha researcher with systematic academic analysis to address critical digital safety challenges.

AI安全00后语言内容审核青少年保护

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。