arXiv:2602.00770cs.CL2026-02

揭示大模型推理时内部状态的动态演化机制

Reasoning as State Transition: A Representational Analysis of Reasoning Evolution in Large Language Models

  • 从表征视角分析模型推理过程中的内部状态变化
  • 推理生成时表征分布持续迁移,非静态不变
  • 训练提升的是状态转移能力,而非初始表征质量

大型语言模型在推理任务上表现卓越,促使研究者关注其推理能力在训练过程中的演化。以往工作多依赖生成结果分析,将推理过程视为黑箱,掩盖了内部变化。本文引入表征视角,探究模型内部状态的动态演变。跨不同训练阶段的实验发现,训练后仅带来有限的初始表征质量提升。更重要的是,与非推理任务不同,推理过程中存在显著的连续表征分布偏移。对比分析表明,训练使模型具备驱动状态向更优分布转移的能力。统计分析确认生成正确性与最终表征高度相关;反事实实验指出,生成词元语义是状态转移的主要驱动力,而非推理时额外计算或参数差异。本研究为理解推理机制及训练对推理能力的增强提供了新视角。

原文摘要 · Abstract (English)

Large Language Models have achieved remarkable performance on reasoning tasks, motivating research into how this ability evolves during training. Prior work has primarily analyzed this evolution via explicit generation outcomes, treating the reasoning process as a black box and obscuring internal changes. To address this opacity, we introduce a representational perspective to investigate the dynamics of the model's internal states. Through comprehensive experiments across models at various training stages, we discover that post-training yields only limited improvement in static initial representation quality. Furthermore, we reveal that, distinct from non-reasoning tasks, reasoning involves a significant continuous distributional shift in representations during generation. Comparative analysis indicates that post-training empowers models to drive this transition toward a better distribution for task solving. To clarify the relationship between internal states and external outputs, statistical analysis confirms a high correlation between generation correctness and the final representations; while counterfactual experiments identify the semantics of the generated tokens, rather than additional computation during inference or intrinsic parameter differences, as the dominant driver of the transition. Collectively, we offer a novel understanding of the reasoning process and the effect of training on reasoning enhancement, providing valuable insights for future model analysis and optimization.

大模型推理机制表征演化

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。