arXiv:2502.12576cs.CLcs.AI2025-02中稿 · publication in the…被引 1

用模糊理论评估句编码器在诱骗风险分类中的表现,发现模型对隐晦语言敏感度不足。

A Fuzzy Evaluation of Sentence Encoders on Grooming Risk Classification

  • 采用模糊理论将人类对诱骗行为的判断转化为风险等级评分
  • 模型在含大量生僻词(OOV)的间接表达样本中误判率显著升高
  • 适用于需要理解隐蔽网络诱骗行为的执法与心理干预研究

随着社交媒体发展,儿童在网络环境中面临诱骗风险日益加剧。在线对话中的诱骗行为往往不直接涉及性内容,施害者通过逐步建立信任关系并使用隐晦、编码的语言逃避检测。尽管已有研究微调Transformer模型自动识别聊天中的诱骗行为,但忽视了编码语言对模型预测的影响及其与人类感知的一致性。本文针对三种参与者群体(执法部门、真实受害者、诱饵人员),评估双编码器在区分不同诱骗风险等级任务中的表现。基于模糊理论框架,将人类对诱骗行为的主观判断映射为实际风险程度。分析显示,微调模型难以识别使用间接言语路径和编码语言的案例;这些案例普遍存在更高比例的未登录词(OOV),导致模型误判。研究强调需构建更鲁棒的模型以从噪声聊天数据中识别隐蔽的诱骗语言。

原文摘要 · Abstract (English)

With the advent of social media, children are becoming increasingly vulnerable to the risk of grooming in online settings. Detecting grooming instances in an online conversation poses a significant challenge as the interactions are not necessarily sexually explicit, since the predators take time to build trust and a relationship with their victim. Moreover, predators evade detection using indirect and coded language. While previous studies have fine-tuned Transformers to automatically identify grooming in chat conversations, they overlook the impact of coded and indirect language on model predictions, and how these align with human perceptions of grooming. In this paper, we address this gap and evaluate bi-encoders on the task of classifying different degrees of grooming risk in chat contexts, for three different participant groups, i.e. law enforcement officers, real victims, and decoys. Using a fuzzy-theoretic framework, we map human assessments of grooming behaviors to estimate the actual degree of grooming risk. Our analysis reveals that fine-tuned models fail to tag instances where the predator uses indirect speech pathways and coded language to evade detection. Further, we find that such instances are characterized by a higher presence of out-of-vocabulary (OOV) words in samples, causing the model to misclassify. Our findings highlight the need for more robust models to identify coded language from noisy chat inputs in grooming contexts.

诱骗检测模糊理论句编码器网络安全

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。