揭露学术论文中未声明的AI写作现象,揭示其广泛存在且难被发现。
Academ-AI: documenting the undisclosed use of generative artificial intelligence in academic publishing
- 通过识别大模型特有的异常语调,发现768例未声明的AI使用痕迹
- 高影响因子期刊与高APC论文更易出现未声明AI使用,反向凸显监管漏洞
- 多数问题未被纠正,建议出版商强化可检测案例的政策执行
自生成式AI工具(如OpenAI的ChatGPT)普及以来,研究者在论文撰写中使用该技术的情况日益普遍。学术出版界共识是此类使用必须在发表文章中明确披露。本研究通过分析前768个疑似未声明使用AI的案例,发现该问题普遍存在,已渗透至知名出版社的期刊、会议论文集及教科书中。未声明的AI使用多出现在引用率高、文章处理费(APC)高的期刊,这些本应具备更强监管能力的出版物反而更易忽视此问题。极少数案例在发表后被修正,且纠正措施往往不足以解决问题。所分析的768例可能仅为实际数量的一小部分,大量未声明使用可能无法被检测。出版商应对可识别的案例严格执行禁止未声明使用AI的政策,这是当前学术出版界遏制隐蔽AI泛滥最有效的防御手段。此为先前预印本的更新版本。
原文摘要 · Abstract (English)
Since generative artificial intelligence (AI) tools such as OpenAI's ChatGPT became widely available, researchers have used them in the writing process. The consensus of the academic publishing community is that such usage must be declared in the published article. Academ-AI documents examples of suspected undeclared AI usage in the academic literature, discernible primarily due to the appearance in research papers of idiosyncratic verbiage characteristic of large language model (LLM)-based chatbots. This analysis of the first 768 examples collected reveals that the problem is widespread, penetrating the journals, conference proceedings, and textbooks of highly respected publishers. Undeclared AI seems to appear in journals with higher citation metrics and higher article processing charges (APCs), precisely those outlets that should theoretically have the resources and expertise to avoid such oversights. An extremely small minority of cases are corrected post publication, and the corrections are often insufficient to rectify the problem. The 768 examples analyzed here likely represent a small fraction of the undeclared AI present in the academic literature, much of which may be undetectable. Publishers must enforce their policies against undeclared AI usage in cases that are detectable; this is the best defense currently available to the academic publishing community against the proliferation of undisclosed AI. This is an updated version of a previous preprint.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。