统一英语指代消解标注标准,释放跨数据集评测资源
Unifying the Scope of Bridging Anaphora Types in English: Bridging Annotations in ARRAU and GUM
- 对比GUM、GENTLE和ARRAU三数据集的指代标注规范
- 发现不同数据集对指代现象的定义差异显著
- 公开可直接使用的标准化测试集,支持跨领域评测
跨核心参考资源的指代标注比较困难,主要源于定义与标注模式缺乏统一,且各资源覆盖文本领域有限。为缓解领域覆盖不足并整合标注体系,本文对比了GUM、GENTLE和ARRAU语料库的标注指南,并借助可解释性预测模型分析其中的指代实例。研究发现,不同数据集对指代现象的标注类型存在显著差异。基于此,我们发布了经过统一处理、细粒度分类的GUM、GENTLE及ARRAU《华尔街日报》数据集测试集,以推动跨领域指代消解任务的可比、可靠评估。
原文摘要 · Abstract (English)
Comparing bridging annotations across coreference resources is difficult, largely due to a lack of standardization across definitions and annotation schemas and narrow coverage of disparate text domains across resources. To alleviate domain coverage issues and consolidate schemas, we compare guidelines and use interpretable predictive models to examine the bridging instances annotated in the GUM, GENTLE and ARRAU corpora. Examining these cases, we find that there is a large difference in types of phenomena annotated as bridging. Beyond theoretical results, we release a harmonized, subcategorized version of the test sets of GUM, GENTLE and the ARRAU Wall Street Journal data to promote meaningful and reliable evaluation of bridging resolution across domains.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。