arXiv:2503.00444cs.CL2025-03被引 7

开源意大利语隐喻数据库,含996个隐喻的多维度评测数据。

Figurative Archive: an open dataset and web-based application for the study of metaphor

  • 整合11项研究的隐喻材料,构建996个意大利语隐喻数据集。
  • 提供熟悉度、语义距离、解释偏好等20余项量化指标。
  • 支持自定义查询,适合语言认知与计算模型研究者使用。

近年来,隐喻研究持续增长,因其揭示了语言与认知过程的重要窗口。同时,对严谨构建且广泛规范的实验材料需求也日益增加。本文介绍图式档案(Figurative Archive),一个包含996个意大利语隐喻的开放数据库,涵盖日常与文学隐喻,结构与语义领域多样。该数据库基于11项研究中使用的刺激材料构建,提供了从熟悉度到语义距离、首选解释等多项评分与语料库度量,并通过熟悉度与其他指标间的相关性进行了验证。其创新之处在于规模更大、引入隐喻包容性度量以符合非歧视性语言使用建议,并配有支持自定义查询的网页界面。本文还提供使用指南,供研究隐喻处理及人类与计算模型间隐喻特征关系的研究者参考。

原文摘要 · Abstract (English)

Research on metaphor has steadily increased over the last decades, as this phenomenon opens a window into a range of linguistic and cognitive processes. At the same time, the demand for rigorously constructed and extensively normed experimental materials increased as well. Here, we present the Figurative Archive, an open database of 996 metaphors in Italian enriched with rating and corpus-based measures (from familiarity to semantic distance and preferred interpretations), derived by collecting stimuli used across 11 studies. It includes both everyday and literary metaphors, varying in structure and semantic domains, and is validated based on correlations between familiarity and other measures. The Archive has several aspects of novelty: it is increased in size compared to previous resources; it offers a measure of metaphor inclusiveness, to comply with recommendations for non-discriminatory language use; it is displayed in a web-based interface, with features for a customized consultation. We provide guidelines for using the Archive to source materials for studies investigating metaphor processing and relationships between metaphor features in humans and computational models.

隐喻研究语言认知开放数据意大利语

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。