arXiv:2506.08172cs.CL2025-06被引 1

用文学理论构建评估AI微小说的客观标准

Can Artificial Intelligence Write Like Borges? An Evaluation Protocol for Spanish Microfiction

  • 基于文学理论设计多维度评估框架
  • 专家与爱好者验证协议有效性和可靠性
  • 适合评估AI生成作品的文学价值

自动故事生成已研究超过60年,大语言模型可生成叙事连贯、语言通顺的短篇小说。然而,针对其文学价值——特别是美学品质——的严格评估仍鲜有关注。本文提出GrAImes:一种基于文学理论的评估协议,涵盖主题一致性、文本清晰度、阐释深度和美学质量等维度,为评估AI生成的微小说提供客观框架。通过文学专家与爱好者对协议的验证,结果表明该方法具备可靠性和实用性。该协议将为未来评估自动化创作的文学价值奠定基础。

原文摘要 · Abstract (English)

Automated story writing has been a subject of study for over 60 years. Large language models can generate narratively consistent and linguistically coherent short fiction texts. Despite these advancements, rigorous assessment of such outputs for literary merit - especially concerning aesthetic qualities - has received scant attention. In this paper, we address the challenge of evaluating AI-generated microfictions and argue that this task requires consideration of literary criteria across various aspects of the text, such as thematic coherence, textual clarity, interpretive depth, and aesthetic quality. To facilitate this, we present GrAImes: an evaluation protocol grounded in literary theory, specifically drawing from a literary perspective, to offer an objective framework for assessing AI-generated microfiction. Furthermore, we report the results of our validation of the evaluation protocol, as answered by both literature experts and literary enthusiasts. This protocol will serve as a foundation for evaluating automatically generated microfictions and assessing their literary value.

文本生成文学评估AI艺术

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。