跨语言仇恨言论检测与反言辞生成的全面指南
Multilingual Hate Speech Detection and Counterspeech Generation: A Comprehensive Survey and Practical Guide
- 提出三阶段框架:任务设计、数据整理、评估
- 指出低资源语言数据稀缺与文化表达遗漏问题
- 适合研究者、开发者及政策制定者参考实践
在多语言网络环境中对抗仇恨言论,需超越以英语为中心的模型,捕捉全球在线话语的文化与语言多样性。本文系统综述并提供多语言仇恨言论检测与反言辞生成的实践指南,整合自然语言处理最新进展。分析单语系统在非英语及混合语言情境中失效的原因,常忽略隐含仇恨和文化特有表达。为此,提出结构化三阶段框架——任务设计、数据整理、评估,基于前沿数据集、模型与指标。综述整合多语言资源与技术进展,揭示持续挑战:低资源语言数据匮乏、系统开发中的公平性与偏见、多模态解决方案需求。通过融合技术进步与伦理文化考量,为研究者、从业者及政策制定者提供可扩展的指导,助力构建情境感知、包容性强的系统。本路线图推动通过更公平、高效的跨语言检测与反言辞生成提升在线安全。
原文摘要 · Abstract (English)
Combating online hate speech in multilingual settings requires approaches that go beyond English-centric models and capture the cultural and linguistic diversity of global online discourse. This paper presents a comprehensive survey and practical guide to multilingual hate speech detection and counterspeech generation, integrating recent advances in natural language processing. We analyze why monolingual systems often fail in non-English and code-mixed contexts, missing implicit hate and culturally specific expressions. To address these challenges, we outline a structured three-phase framework - task design, data curation, and evaluation - drawing on state-of-the-art datasets, models, and metrics. The survey consolidates progress in multilingual resources and techniques while highlighting persistent obstacles, including data scarcity in low-resource languages, fairness and bias in system development, and the need for multimodal solutions. By bridging technical progress with ethical and cultural considerations, we provide researchers, practitioners, and policymakers with scalable guidelines for building context-aware, inclusive systems. Our roadmap contributes to advancing online safety through fairer, more effective detection and counterspeech generation across diverse linguistic environments.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。