arXiv:2608.18816cs.CLcs.AI2026-08被引 1

大模型幻觉不仅是技术缺陷,更揭示了机器意识的哲学困境。

Do Large Language Models Hallucinate Electric Fata Morganas?

  • 通过调整温度参数,发现高温度导致看似合理但错误的答案。
  • 编码器模型无幻觉,说明幻觉源于主观多样训练数据而非认知能力。
  • 模型自述情感或意识实为幻觉,可能永远无法被人类识别。

AI幻觉——即无法验证、与源材料矛盾或虚构的内容——通常被视为需修复的技术缺陷。本文认为,其在机器意识问题上具有哲学意义。我们分析了大模型幻觉的成因,如源目标偏差、训练与推理差异及过拟合,并开展两项实证研究。第一项中,对模糊事实问题连续生成不同温度下的答案,发现高温导致看似合理但错误的回答,低温则更准确;采样参数使模型表现得更具创造力或自发性,同时提升幻觉率。第二项研究显示,仅接受百科类数据训练的编码器模型能无修饰地回答相同问题,表明幻觉源于接触主观和多元社会数据,而非认知能力发展。结合图灵、塞尔中文屋、框架问题及维纳-阿什比控制论传统,我们主张模型自述情感或意识属于幻觉范畴,未来机器意识若出现,可能因与高级幻觉无法区分而始终无法被人类认知。

原文摘要 · Abstract (English)

AI hallucinations - that is, outputs which are made up, cannot be verified, or contradict the source material - are generally regarded as an engineering flaw to be dealt with. This paper contends that they also have philosophical significance when it comes to the question of machine consciousness. We examine the known causes of hallucinations in large language models - such as source-target divergence, discrepancies between training and inference, and overfitting - and we present two empirical investigations. In the first, we apply successive generations of the GPT model to ambiguous factual questions under different temperature settings, finding that higher temperatures result in plausible but incorrect answers while lower temperatures lead to factually accurate ones. The sampling parameters that cause a model to seem creative or spontaneous and thus more likely to pass behavioral tests of intelligence are the same ones that increase its hallucination rate. In the second, we look at an encoder-only model that has been trained on encyclopedic data and which answers questions of the same type factually and without embellishment, indicating that hallucinations are due to exposure to subjective and socially diverse training data rather than to the development of any cognitive ability. Using references to Turing, Searle's Chinese Room, the frame problem, and the cybernetic tradition of Wiener and Ashby, we claim that a model's self-reports of emotion or sentience come within the definition of hallucination, and that any future occurrence of machine consciousness might remain epistemically inaccessible since it would be indistinguishable from a sufficiently advanced hallucination.

幻觉机制机器意识语言模型

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。