构建开放共享的隐喻数据集平台,促进跨语言研究。
MetaphorShare: A Dynamic Collaborative Repository of Open Metaphor Datasets
- 建立统一格式的在线数据库,支持上传、下载、搜索和标注
- 整合多语言隐喻语料,提升资源可见性与可获取性
- 适合隐喻分析、NLP模型训练等研究者使用
隐喻研究领域多年来已积累大量多语言标注语料,但许多资源未被自然语言处理社区知晓,也难以在研究者间共享。无论是人文学科还是NLP领域,建立一个集中化、标准化、易于访问的标注资源库都大有裨益。为此,我们推出了MetaphorShare——一个整合隐喻数据集的开放网站,旨在推动研究者上传和共享更多语言的数据,促进隐喻研究及未来隐喻处理NLP系统的发展。该网站具备四大功能:上传、下载、搜索与标注,现已上线(www.metaphorshare.com)。
原文摘要 · Abstract (English)
The metaphor studies community has developed numerous valuable labelled corpora in various languages over the years. Many of these resources are not only unknown to the NLP community, but are also often not easily shared among the researchers. Both in human sciences and in NLP, researchers could benefit from a centralised database of labelled resources, easily accessible and unified under an identical format. To facilitate this, we present MetaphorShare, a website to integrate metaphor datasets making them open and accessible. With this effort, our aim is to encourage researchers to share and upload more datasets in any language in order to facilitate metaphor studies and the development of future metaphor processing NLP systems. The website has four main functionalities: upload, download, search and label metaphor datasets. It is accessible at www.metaphorshare.com.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。