arXiv:2410.01677cs.AI2024-10被引 7

用文字乱序测试大模型思维,发现其认知模式像人但机制不同。

Mind Scramble: Unveiling Large Language Model Psychology Via Typoglycemia

  • 通过文字乱序实验模拟人类心理,研究大模型的认知行为。
  • 模型在乱序文本中准确率下降,耗时增加,复杂任务受影响更明显。
  • 不同模型有独特认知模式,可作为无新数据的评估基准。

大语言模型(LLMs)在外部行为和内部机制方面的研究,为解决物理世界中的复杂任务带来了希望。研究表明,如GPT-4等强大模型正展现出类似人类的规划、推理与反思能力。本文提出“大模型心理学”这一新研究方向,借鉴人类心理实验方法探究模型的认知行为与机制。我们将心理学中的“文字乱序”现象迁移至大模型研究,发现:(I)大模型在宏观层面表现出类人行为,如任务准确率下降、令牌消耗和时间增加;(II)不同模型对乱序输入的鲁棒性各异,使文字乱序成为无需新数据的模型评估基准;(III)复杂逻辑任务(如数学)在乱序形式下更难处理;(IV)每种模型在不同任务中呈现独特且一致的“认知模式”,揭示其内在心理过程。通过对隐藏层的深入分析,我们为未来大模型心理学与可解释性研究提供了路径。

原文摘要 · Abstract (English)

Research into the external behaviors and internal mechanisms of large language models (LLMs) has shown promise in addressing complex tasks in the physical world. Studies suggest that powerful LLMs, like GPT-4, are beginning to exhibit human-like cognitive abilities, including planning, reasoning, and reflection. In this paper, we introduce a research line and methodology called LLM Psychology, leveraging human psychology experiments to investigate the cognitive behaviors and mechanisms of LLMs. We migrate the Typoglycemia phenomenon from psychology to explore the "mind" of LLMs. Unlike human brains, which rely on context and word patterns to comprehend scrambled text, LLMs use distinct encoding and decoding processes. Through Typoglycemia experiments at the character, word, and sentence levels, we observe: (I) LLMs demonstrate human-like behaviors on a macro scale, such as lower task accuracy and higher token/time consumption; (II) LLMs exhibit varying robustness to scrambled input, making Typoglycemia a benchmark for model evaluation without new datasets; (III) Different task types have varying impacts, with complex logical tasks (e.g., math) being more challenging in scrambled form; (IV) Each LLM has a unique and consistent "cognitive pattern" across tasks, revealing general mechanisms in its psychology process. We provide an in-depth analysis of hidden layers to explain these phenomena, paving the way for future research in LLM Psychology and deeper interpretability.

大模型心理认知行为可解释性

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。