AI写作检测模型误判自闭症写作风格,可能造成不公平歧视。
The Misclassification of Autistic Writing as AI-Generated
- 用6万条Reddit帖子对比自闭症倾向作者与普通用户文本。
- 自闭症倾向文本被误标为AI生成的比例显著更高。
- 模型存在对特殊群体的隐性偏见,需重新审视其伦理应用。
近期研究发现,人工智能(AI)文本检测模型无法准确识别AI生成内容,且可能对少数群体产生偏见。本研究实证检验了关于自闭症写作者更常被误判为使用AI生成文本的传闻。分析约6万条来自Reddit的帖子,分为‘可能自闭症’和‘普通社区’两个子语料库,对比OpenAI GPT-2检测模型输出的概率分布。结果显示,两组中被标记为AI生成的文本均不足2%,但自闭症倾向组被标记比例显著更高。尽管两类文本在特征上存在差异,但与已知AI生成文本特征的关联并不直接。该结果表明,当前广泛使用的AI检测模型可能对自闭症写作者存在系统性偏见,亟需对其伦理使用进行批判性审查,尤其在学术场景中应谨慎对待。
原文摘要 · Abstract (English)
Recent findings suggest that detection models for artificial intelligence (AI) cannot accurately identify AI-generated text and may exhibit bias against certain minority groups. In the present study, anecdotal claims that autistic writers more often have their work flagged as AI-generated are examined empirically. A corpus of approximately 60,000 Reddit posts split into "likely-autistic" and "general-Reddit" subcorpora is used to compare the distribution of probabilities output by the OpenAI GPT-2 detection model. Differences in textual features between subcorpora are observed and compared to reported features of AI-generated text. Results showed that while less than two-percent of either subcorpus was flagged as AI-generated by the model, significantly more texts from the likely-autistic subcorpus were flagged. Connections between features of text with likely-autistic authors and AI-generated text were not straightforward. The widespread use of AI-detection models with a potential bias against autistic writers in their output prompts ethical scrutiny, and the authors recommend further critical examination of the models themselves as well as their use in academic contexts.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。