测试大模型对汉语话题结构的语法敏感度,聚焦岛屿约束问题。
Evaluating LLMs on Chinese Topic Constructions: A Research Proposal Inspired by Tian et al. (2024)
- 基于田等(2024)设计实验框架,评估模型对汉语句法的掌握。
- 暂未开展实验,提出可验证的测试方案与评估路径。
- 适合研究中文语言模型语法能力的学者参考。
本文提出一个评估大型语言模型(LLMs)在汉语话题结构方面表现的框架,重点关注其对岛屿约束的敏感性。受田等(2024)启发,本文勾勒了测试大模型汉语句法知识的实验设计。尽管尚未开展实际实验,该提案旨在为未来研究奠定基础,并邀请学术界对方法论提出反馈。
原文摘要 · Abstract (English)
This paper proposes a framework for evaluating large language models (LLMs) on Chinese topic constructions, focusing on their sensitivity to island constraints. Drawing inspiration from Tian et al. (2024), we outline an experimental design for testing LLMs' grammatical knowledge of Mandarin syntax. While no experiments have been conducted yet, this proposal aims to provide a foundation for future studies and invites feedback on the methodology.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。