arXiv:2601.18353cs.AIcs.CL2026-01被引 10

用高质量书籍微调后,AI写作水平超越人类专家。

Can Good Writing Be Generative? Expert-Level AI Writing Emerges through Fine-Tuning on High-Quality Books

  • 用经典著作微调大模型,让AI模仿作家风格
  • 专家对微调后AI写作的偏好从38%升至62%
  • 引发创作者身份危机,挑战艺术创作本质

创意写作长期被视为人类独有能力,依赖无法被机器复制的个人风格。然而,生成式AI可在几秒内模仿数千种写作风格,边际成本几乎为零。为深入理解这一现象,我们组织了一场行为实验:28位文学硕士(MFA)作家与三款大型语言模型在模仿50位受好评作家方面展开竞争。通过28位专家评委和131位普通评委进行盲测对比,结果显示,在仅使用上下文提示时,专家更偏好人写作(82.7%),但经过对作者全部作品的微调后,偏好转向AI写作(62%)。普通评委则始终更青睐AI。专家访谈揭示,他们对AI写作的接受引发了身份认同危机,削弱了审美自信,动摇了对“好写作”的定义。研究挑战了关于AI创作局限性的主流话语,也引发了对创意劳动未来的根本性思考。

原文摘要 · Abstract (English)

Creative writing has long been considered a uniquely human endeavor, requiring voice and style that machines could not replicate. This assumption is challenged by Generative AI that can emulate thousands of author styles in seconds with negligible marginal labor. To understand this better, we conducted a behavioral experiment where 28 MFA writers (experts) competed against three LLMs in emulating 50 critically acclaimed authors. Based on blind pairwise comparisons by 28 expert judges and 131 lay judges, we find that experts preferred human writing in 82.7% of cases under the in-context prompting condition but this reversed to 62% preference for AI after fine-tuning on authors' complete works. Lay judges, however, consistently preferred AI writing. Debrief interviews with expert writers revealed that their preference for AI writing triggered an identity crisis, eroding aesthetic confidence and questioning what constitutes "good writing." These findings challenge discourse about AI's creative limitations and raise fundamental questions about the future of creative labor.

AI写作创作伦理大模型微调

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。