arXiv:2510.02869cs.CYcs.AI2025-10

机器能识别美吗?研究发现美的图像让不同模型的表征更相似。

Representing Beauty: Towards a Participatory but Objective Latent Aesthetics

  • 用跨模型表征一致性检验美学判断,发现美图引发模型表征趋同
  • 美图使不同训练数据与模态的模型产生更对齐的表示,非美图则不然
  • 强调人类感知与创作在塑造模型美学空间中的核心作用

机器如何理解美?尽管美是文化与经验交织却哲学难解的概念,深度学习系统正逐步具备建模审美判断的能力。本文探讨神经网络在面对形式多样对象时,仍能表征‘美’的能力。基于跨模型表征收敛的研究,我们发现:经过不同数据与模态训练的模型,其对美图的表征更相似、更对齐,而非美图则无此现象。这一结果暗示美的形式结构具有现实基础,而不仅是社会建构的反映。我们认为,这种现实性源于美学形式在物理与文化双重实质上的共同奠基。人类感知与创造行为在塑造深度学习模型的潜在空间中起关键作用,但机器并非仅模仿创作,而是能从规模视角产生新颖创意洞察。研究显示,人机协同创造不仅是可能的,更是根本性的——美既是文化生产的终极目标,也是机器感知的引力中心。

原文摘要 · Abstract (English)

What does it mean for a machine to recognize beauty? While beauty remains a culturally and experientially compelling but philosophically elusive concept, deep learning systems increasingly appear capable of modeling aesthetic judgment. In this paper, we explore the capacity of neural networks to represent beauty despite the immense formal diversity of objects for which the term applies. By drawing on recent work on cross-model representational convergence, we show how aesthetic content produces more similar and aligned representations between models which have been trained on distinct data and modalities - while unaesthetic images do not produce more aligned representations. This finding implies that the formal structure of beautiful images has a realist basis - rather than only as a reflection of socially constructed values. Furthermore, we propose that these realist representations exist because of a joint grounding of aesthetic form in physical and cultural substance. We argue that human perceptual and creative acts play a central role in shaping these the latent spaces of deep learning systems, but that a realist basis for aesthetics shows that machines are not mere creative parrots but can produce novel creative insights from the unique vantage point of scale. Our findings suggest that human-machine co-creation is not merely possible, but foundational - with beauty serving as a teleological attractor in both cultural production and machine perception.

美学表征深度学习人机协同

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。