研究听觉注意力如何影响多人对话场景中的感知,发现目标导向注意力最有效。
Perception of dynamic multi-speaker auditory scenes under different modes of attention
- 通过空间、说话人和整体三种注意模式对比实验
- 对象注意比特征注意和全局注意更易察觉语音变调
- 神经机制不同,底向注意在追踪中作用减弱
注意力并非单一形式,而是以多种方式促进高效认知处理。在听觉领域,注意力可优先处理听觉场景中的相关声音,既可通过场景元素自下而上吸引,也可通过特征、物体或整个场景自上而下引导。这些注意模式如何交互及其神经基础是否不同尚不明确。本研究在受控的“鸡尾酒会”范式中,让被试聆听相同刺激,分别关注空间位置(特征基)、说话人(对象基)或整个场景(全局或自由聆听),同时检测语音音高偏差。结果表明,对象注意在感知上优于特征注意或全局注意。此外,对象注意与空间注意激活不同的神经机制,且对自下而上显著性反应不同。值得注意的是,尽管自下而上显著性有助于初始分离听觉对象,但一旦注意力被主动分配,其作用便减弱。解码脑电数据还发现,在全局注意模式中存在一种源采样机制,而在对象或空间注意模式中不存在。总体而言,同一声学场景的感知差异取决于听觉任务,由自上而下与自下而上过程相互作用决定。
原文摘要 · Abstract (English)
Attention is not monolithic; rather, it operates in multiple forms to facilitate efficient cognitive processing. In the auditory domain, attention enables the prioritization of relevant sounds in an auditory scene and can be either attracted by elements in the scene in a bottom-up fashion or directed towards features, objects, or the entire scene in a top-down fashion. How these modes of attention interact and whether their neural underpinnings are distinct remains unclear. In this work, we investigate the perceptual and neural correlates of different attentional modes in a controlled "cocktail party" paradigm, where listeners listen to the same stimuli and attend to either a spatial location (feature-based), a speaker (object-based), or the entire scene (global or free-listening) while detecting deviations in pitch of a voice in the scene. Our findings indicate that object-based attention is more perceptually effective than feature-based or global attention. Furthermore, object-based and spatial-based attention engage distinct neural mechanisms and are differentially modulated by bottom-up salience. Notably, while bottom-up salience aids in the initial segregation of auditory objects, it plays a reduced role in object tracking once attention has been voluntarily allocated. In addition, decoding the stimulus envelope from the EEG data revealed a source-sampling scheme in the global attention mode that is not present in the object or spatial modes. Overall, the study shows that the perception of the same acoustic scene differs according to the listening task, guided by an interaction between top-down and bottom-up processes.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。