arXiv:2505.17704cs.CL2025-05

构建首个公开语义草图语料库,探索机器处理方法

SemSketches-2021: experimenting with the machine processing of the pilot semantic sketches corpus

  • 构建开放语义草图语料库并组织共享任务
  • 通过上下文匹配任务验证草图的机器可处理性
  • 适合自然语言处理与语义表示研究者参考

本文探讨了语义草图的机器处理方法,发布了首个开放的语义草图试点语料库。论文讨论了草图构建的多个方面及其可解决的任务类型,并重点开发了针对该语料库的机器处理工具。为此,组织了SemSketches-2021共享任务,参赛者需在匿名草图和包含必要谓词的上下文集合中,将合适的上下文匹配到对应草图上,以评估语义草图的可计算性与应用潜力。

原文摘要 · Abstract (English)

The paper deals with elaborating different approaches to the machine processing of semantic sketches. It presents the pilot open corpus of semantic sketches. Different aspects of creating the sketches are discussed, as well as the tasks that the sketches can help to solve. Special attention is paid to the creation of the machine processing tools for the corpus. For this purpose, the SemSketches-2021 Shared Task was organized. The participants were given the anonymous sketches and a set of contexts containing the necessary predicates. During the Task, one had to assign the proper contexts to the corresponding sketches.

语义草图自然语言处理共享任务

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。