arXiv:2510.18629cs.CL2025-10中稿 · publication in JAS…被引 3

用超声舌动数据可可靠估计发音动力学参数,与电磁数据一致。

Dynamical model parameters from ultrasound tongue kinematics

  • 通过超声影像建模舌部运动动力学,采用线性谐振子模型拟合
  • 超声与电磁数据估算的动态参数高度一致,误差在可接受范围
  • 适合语音控制研究者,尤其关注无创舌动测量的团队

发音控制可建模为动力系统,其中发音器官被驱动至目标位置。这类模型通常使用软组织点数据(如电磁言语仪, EMA)进行评估,但超声成像技术的新进展使其成为有前景的替代方案。本文评估了能否从超声舌动数据中可靠估计线性谐振子模型的参数,并与同步采集的EMA数据结果进行对比。结果显示,超声与EMA获得的动态参数具有可比性;同时,下颌短肌腱追踪也能有效捕捉下颌运动。这些结果支持利用超声舌动数据评估发音动力学模型。

原文摘要 · Abstract (English)

The control of speech can be modelled as a dynamical system in which articulators are driven toward target positions. These models are typically evaluated using fleshpoint data, such as electromagnetic articulography (EMA), but recent methodological advances make ultrasound imaging a promising alternative. We evaluate whether the parameters of a linear harmonic oscillator can be reliably estimated from ultrasound tongue kinematics and compare these with parameters estimated from simultaneously-recorded EMA data. We find that ultrasound and EMA yield comparable dynamical parameters, while mandibular short tendon tracking also adequately captures jaw motion. This supports using ultrasound kinematics to evaluate dynamical articulatory models.

语音建模超声成像动力系统

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。