arXiv:2502.11066cs.CL2025-02EMNLP被引 3

通过正则化与信息对齐提升大模型的组合推理能力

CARMA: Enhanced Compositionality in LLMs via Advanced Regularisation and Mutual Information Alignment

  • 引入互信息正则与层间稳定性约束,防止特征碎片化
  • 减少微调带来的表示变异,稳定词元表征
  • 适配各类模型架构,适合需可靠组合推理的场景

大语言模型在组合泛化方面存在困难,难以系统性地组合已学成分来理解新输入。尽管架构修改、微调和数据增强能部分改善组合性,但常受限于适应性差、可扩展性低或真实数据上收益递减。为此,我们提出CARMA,一种增强大模型组合推理稳定性与鲁棒性的干预方法,同时保持微调性能。CARMA采用互信息正则化与逐层稳定性约束,缓解特征碎片化,确保层间与层内表征结构一致。我们在反向词典建模与情感分类任务上评估,衡量语义一致性、性能稳定性及对词汇扰动的鲁棒性。结果表明,CARMA降低了微调引入的变异性,稳定了词元表示,并提升了组合推理能力。尽管效果因模型架构而异,其核心优势在于强化已有学习结构而非添加新能力,是一种可扩展的辅助方法。这表明,将CARMA与微调结合可在保持任务性能的同时,显著提升大模型的组合泛化能力。

原文摘要 · Abstract (English)

Large language models (LLMs) struggle with compositional generalisation, limiting their ability to systematically combine learned components to interpret novel inputs. While architectural modifications, fine-tuning, and data augmentation improve compositionality, they often have limited adaptability, face scalability constraints, or yield diminishing returns on real data. To address this, we propose CARMA, an intervention that enhances the stability and robustness of compositional reasoning in LLMs while preserving fine-tuned performance. CARMA employs mutual information regularisation and layer-wise stability constraints to mitigate feature fragmentation, ensuring structured representations persist across and within layers. We evaluate CARMA on inverse dictionary modelling and sentiment classification, measuring its impact on semantic consistency, performance stability, and robustness to lexical perturbations. Results show that CARMA reduces the variability introduced by fine-tuning, stabilises token representations, and improves compositional reasoning. While its effectiveness varies across architectures, CARMA's key strength lies in reinforcing learned structures rather than introducing new capabilities, making it a scalable auxiliary method. These findings suggest that integrating CARMA with fine-tuning can improve compositional generalisation while maintaining task-specific performance in LLMs.

大模型组合性正则化推理

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。