研究自发对话中声音与面部同步如何影响沟通成功感知。
Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

- 从视频通话中提取语音、停顿、面部动作等多模态特征
- 发现同步行为与更高沟通成功率显著相关
- 适合人机交互和远程沟通优化研究者参考
人们常无意识地调整说话方式以匹配对方,这种现象与互动投入和关系建立相关。尽管在任务导向对话中已有研究,但自然情境下的非任务性及虚拟场景中的同步现象仍不明确。本研究分析大规模自发双人视频通话数据,考察对话动态与感知互动质量的关系。提取了包含话轮转换、停顿、面部运动及音高、强度等声学特征的多模态信号。通过对话后评分的因子分析量化感知沟通成功度。结果表明,自发对话中可稳定检测到同步行为,且与更高感知成功率相关。研究识别出关键互动标志,为提升有效沟通提供了干预方向。
原文摘要 · Abstract (English)
Individuals often align their speaking patterns with their interlocutors, a phenomenon linked to engagement and rapport. While well documented in task-oriented dialogues, less is known about entrainment in naturalistic, non-task and virtual settings. In this study, we analyze a large corpus of spontaneous dyadic Zoom conversations to examine how conversational dynamics relate to perceived interaction quality. We extract multimodal features encompassing turn-taking, pauses, facial movements, and acoustic measures such as pitch and intensity. Perceived conversational success was quantified via factor analysis of post-conversation ratings. Results demonstrate that entrainment reliably detected in spontaneous speech and correlates with higher perceived success. These findings identify key interactional markers of conversational quality and highlight opportunities for targeted interventions to foster more effective and engaging communication.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。