AI写的捷克诗歌比人类还难分辨,但读者偏见让机器作品被低估。
The author is dead, but what if they never lived? A reception experiment on Czech AI- and human-authored poetry
- 用捷克语诗歌测试人类识别AI与真人作品的能力。
- 识别准确率仅45.8%,接近随机水平,说明生成效果逼真。
- 即便AI诗更受欢迎,人们因以为是机器所作而评分更低。
大型语言模型在生成创意文本方面能力日益增强,但多数关于AI诗歌的研究集中于英语——这正是训练数据主导的语言。本文研究了捷克母语者对AI与人类创作的捷克诗歌的感知差异,考察其是否能识别作者身份及如何进行审美评价。参与者在判断作者身份时表现接近随机(平均正确率45.8%),表明捷克语AI诗歌与真人作品难以区分。审美评估显示强烈作者身份偏见:当参与者认为诗歌为AI所作时,评分显著偏低,尽管实际上AI诗歌的平均评分不劣于甚至优于人类作品。逻辑回归模型发现,诗歌越受喜爱,越难准确判断作者身份。诗歌熟悉度或文学背景对识别准确率无影响。结果表明,即使在形态复杂、训练数据稀缺的斯拉夫语系语言如捷克语中,AI也能生成极具说服力的诗歌。研究揭示读者对作者身份的信念与诗歌美学评价之间存在内在关联。
原文摘要 · Abstract (English)
Large language models are increasingly capable of producing creative texts, yet most studies on AI-generated poetry focus on English -- a language that dominates training data. In this paper, we examine the perception of AI- and human-written Czech poetry. We ask if Czech native speakers are able to identify it and how they aesthetically judge it. Participants performed at chance level when guessing authorship (45.8\% correct on average), indicating that Czech AI-generated poems were largely indistinguishable from human-written ones. Aesthetic evaluations revealed a strong authorship bias: when participants believed a poem was AI-generated, they rated it as less favorably, even though AI poems were in fact rated equally or more favorably than human ones on average. The logistic regression model uncovered that the more the people liked a poem, the less probable was that they accurately assign the authorship. Familiarity with poetry or literary background had no effect on recognition accuracy. Our findings show that AI can convincingly produce poetry even in a morphologically complex, low-resource (with respect of the training data of AI models) Slavic language such as Czech. The results suggest that readers' beliefs about authorship and the aesthetic evaluation of the poem are interconnected.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。