用大模型模拟人类语言认知理论,检验其习得机制。
Large Language Models as Proxies for Theories of Human Linguistic Cognition
- 将大模型当作认知理论的代理工具,测试其语言习得能力。
- 验证理论能否从语料中习得特定语言模式,或更易习得类型学常见模式。
- 适合研究语言习得机制与认知建模的学者参考。
本文探讨当前大型语言模型(LLMs)在人类语言认知研究中的潜在作用。聚焦于那些在表征和学习上相对语言中立但与现有大模型存在关键差异的认知理论。通过两类问题展示大模型作为理论代理的潜力:(a) 目标理论能否从给定语料中习得特定语言模式;(b) 目标理论是否使类型学中常见的语言模式比罕见模式更易习得。基于近期文献,我们说明当前大模型可能提供帮助,但指出目前这种帮助仍十分有限。
原文摘要 · Abstract (English)
We consider the possible role of current large language models (LLMs) in the study of human linguistic cognition. We focus on the use of such models as proxies for theories of cognition that are relatively linguistically-neutral in their representations and learning but differ from current LLMs in key ways. We illustrate this potential use of LLMs as proxies for theories of cognition in the context of two kinds of questions: (a) whether the target theory accounts for the acquisition of a given pattern from a given corpus; and (b) whether the target theory makes a given typologically-attested pattern easier to acquire than another, typologically-unattested pattern. For each of the two questions we show, building on recent literature, how current LLMs can potentially be of help, but we note that at present this help is quite limited.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。