检测英瑞议会文本中未披露的AI生成内容,发现2022年起使用量持续上升。
Detecting undisclosed LLM-generated content in parliamentary texts

- 用真实议会文本与AI生成文本训练可解释分类器。
- 2022年后英瑞议会文本中未披露的AI使用率持续上升。
- 适合关注AI透明性与公共治理的研究者和政策制定者。
本文评估了英国与瑞典议会文本中未披露的大型语言模型(LLM)生成内容的普遍程度。在新闻或学术写作中,通常要求明确说明是否使用了如LLM等AI工具;然而,议会文本的相关披露指南较为模糊。为维护透明度与公众信任,建议议员在撰写动议等文本时声明是否使用了AI。为此,我们利用预-LLM时期的议会文本及对应的LLM生成版本,训练了一个可解释的(玻璃箱)文本分类器,并将其应用于包含近期议会文本的测试集。结果显示,自2022年起,两国议会中未披露的LLM使用率呈稳定上升趋势。
原文摘要 · Abstract (English)
In this paper, we evaluate the extent of undisclosed LLM-generated content in texts from the parliaments of the United Kingdom and Sweden. In many areas, such as in journalism or in academic writing, there are often requirements to clearly disclose whether AI tools, such as LLMs, have been used. In the case of parliamentary texts, the guidelines on disclosure of AI use are more vague. However, in order to maintain transparency and retain public trust, it is generally recommended that parliamentarians should state whether or not they have used AI when writing texts, such as parliamentary motions. Here, we train an interpretable (glass-box) text classifier using pre-LLM parliamentary texts and LLM-generated versions of such texts. We then apply the classifier to a test set containing recent parliamentary texts, finding a steady increase in undisclosed LLM use, in both parliaments, from 2022 onwards.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。