arXiv:2501.11012cs.CL2025-01被引 32

评测AI生成英文与多语言文本的检测能力,36团队参与英文任务

GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text Detection: AI vs. Human

  • 构建双语任务:仅英语和多语言机器生成文本二分类检测
  • 36队提交英文任务结果,26队参与多语言任务,提供系统性能排名
  • 分析各模型表现,适合关注生成内容安全的研究者参考

我们介绍了GenAI内容检测任务1——作为COLING 2025 GenAI研讨会的一部分,该共享任务聚焦于二元机器生成文本检测。任务包含单语(英语)和多语言两个子任务。测试阶段共有36支队伍提交了英文任务的正式结果,26支队伍提交了多语言任务的结果。本文全面概述了数据集,总结了结果,包括系统排名与性能分数,详细描述了参赛系统,并对提交结果进行了深入分析。相关资料可访问GitHub仓库:https://github.com/mbzuai-nlp/COLING-2025-Workshop-on-MGT-Detection-Task1

原文摘要 · Abstract (English)

We present the GenAI Content Detection Task~1 -- a shared task on binary machine generated text detection, conducted as a part of the GenAI workshop at COLING 2025. The task consists of two subtasks: Monolingual (English) and Multilingual. The shared task attracted many participants: 36 teams made official submissions to the Monolingual subtask during the test phase and 26 teams -- to the Multilingual. We provide a comprehensive overview of the data, a summary of the results -- including system rankings and performance scores -- detailed descriptions of the participating systems, and an in-depth analysis of submissions. https://github.com/mbzuai-nlp/COLING-2025-Workshop-on-MGT-Detection-Task1

文本检测生成内容AI安全

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。