厘清PESQ多个版本差异,提供最新修正版开源实现
Navigating PESQ: Up-to-Date Versions and Open Implementations
- 梳理二十年来PESQ的版本演变与开源实现
- 指出不同版本间评分差异显著,影响结果可比性
- 开源最新修正版(Corrigendum 2),适合语音质量评估研究者使用
语音感知质量评价(PESQ)虽已被国际电信联盟(ITU)撤销,但仍广泛用于语音质量评估。过去二十年中,PESQ发展出多个版本和公开实现,版本差异显著,令新用户难以选择。本文系统梳理各版本及实现方式,强调必须明确定义使用的具体版本和实现方法,包括多通道信号处理方式,以确保结果可解释性和跨研究比较。同时,我们发布一个包含最新修正(Corrigendum 2)的开源仓库:https://github.com/audiolabs/PESQ,该修正未被其他公开版本支持。
原文摘要 · Abstract (English)
Perceptual Evaluation of Speech Quality (PESQ) is an objective quality measure that remains widely used despite its withdrawal by the International Telecommunication Union (ITU). PESQ has evolved over two decades, with multiple versions and publicly available implementations emerging during this time. Different versions and their updates can be overwhelming, especially for new PESQ users. This work provides practical guidance on the different versions and implementations of PESQ. We show that differences can be significant, especially between PESQ versions. We stress the importance of specifying the exact version and implementation that is used to compute PESQ, and possibly to detail how multi-channel signals are handled. These practices would facilitate the interpretation of results and allow comparisons of PESQ scores between different studies. We also provide a repository that implements the latest corrections to PESQ, i.e., Corrigendum 2, which is not implemented by any other openly available distribution: https://github.com/audiolabs/PESQ.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。