emojis影响说话时的语调,让文字也能传递情感和意图
The Prosody of Emojis
- 通过实验发现说话人会根据表情符号调整语调
- 听众能准确识别出表情符号对应的语义,显著高于随机水平
- 语义差异越大的表情符号,语调差别越大,适合研究人机交互
语调、节奏和语调变化是口语交流的核心特征,传达情绪、意图和话语结构。在缺乏这些线索的文本交流中,表情符号作为视觉替代品,增添了情感与语用层次。本研究通过受控生成任务收集的人类语音数据,直接关联语调与表情符号。利用贝叶斯多层模型分析表明,说话者会系统性地根据表情符号调整语调,且听者能显著高于随机水平地恢复其含义。结果还揭示语调变化存在清晰层级:表情符号语义差异越大,语调差异越明显。这表明表情符号是承载语调意图的重要载体,弥合了数字文本与口语表达之间的鸿沟。
原文摘要 · Abstract (English)
Prosodic features such as pitch, timing, and intonation are central to spoken communication, conveying emotion, intent, and discourse structure. In text-based settings, where these cues are absent, emojis act as visual surrogates that add affective and pragmatic nuance. This study examines how emojis influence prosodic realisation in speech and how listeners interpret prosodic cues to recover emoji meanings. Unlike previous work, we directly link prosody and emojis by analysing human speech data collected through a controlled elicited production task. Using Bayesian multilevel modelling, we show that speakers systematically adapt their prosody based on emoji cues, and that listeners can recover intended meanings significantly above chance. Furthermore, our results reveal a clear hierarchy in prosodic shifts: greater semantic differences between emojis correspond to increased prosodic divergence. These findings suggest that emojis are meaningful carriers of prosodic intent that bridge the gap between digital text and spoken production.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。