音频误导信息具口语与对话特性,传统文本查证失效。
When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms

- 从口语韵律与对话结构重构事实核查机制
- 现有文本查证方法在音频场景准确率显著下降
- 适合语音平台、社会媒体与数字传播研究者
音频平台已超越娱乐功能,成为公众讨论的核心渠道,涵盖播客、广播、WhatsApp语音消息及直播等。数百万节目和数亿听众使音频成为错误信息传播的重要渠道。然而,现有事实核查流程主要针对书面内容,忽视了语音媒体的独特属性。我们指出,音频误导信息不仅是文字的听觉呈现,更因其口语表达(如语调、节奏、情感)和对话结构(跨轮次、多说话人、跨集延续)而具有结构性差异,这些特征带来传统方法难以应对的验证难题。本文综述多模态证据与平台数据,分析现有数据集与方法,揭示传统流程在音频场景中的失效原因。主张推进事实核查需基于语音与对话的真实特性重新设计验证流程。
原文摘要 · Abstract (English)
Audio platforms have evolved beyond entertainment. They have become central to public discourse, from podcasts and radio to WhatsApp voice notes and live streams. With millions of shows and hundreds of millions of listeners, audio platforms are now a major channel for misinformation. Yet existing fact-checking pipelines are mostly designed for written claims, overlooking the unique properties of spoken media. We argue that audio misinformation is not merely textual content with transcripts: it is structurally different because it is both spoken - carrying persuasive force through prosody, pacing, and emotion - and conversational - unfolding across turns, speakers, and episodes. These dual properties introduce verification difficulties that traditional methods rarely face. This position paper synthesizes evidence across modalities and platforms, examines datasets and methods, and highlights why existing pipelines fail on audio. We argue that advancing fact-checking requires rethinking verification pipelines around the spoken and conversational realities of audio.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。