将空管语音与飞行轨迹映射到同一空间,提升繁忙空域态势感知能力。
V2TATC: A Joint Voice-Trajectory Embedding Framework and Dataset for Air Traffic Controller Situational Awareness

- 构建语音与飞行轨迹联合嵌入框架,实现双向查询。
- 在旧金山湾区数据集上验证跨模态检索有效性和一致性。
- 适用于空管辅助系统开发,尤其适合低空通航密集区。
随着国家空域系统内航空交通量持续增长,特别是在低空空域,对可扩展的空中交通管制决策支持工具的需求也日益增加。本文提出语音-轨迹嵌入框架V2TATC,能作为拥堵空域中态势感知的组成部分,协助空管人员实时分析广播式自动相关监视(ADS-B)轨迹或飞行员自然语言表达的意图。该框架表明语音指令与飞行轨迹并非独立,而是共享同一物理参照物——空域中的飞机。V2TATC将语音指令与对应飞机轨迹映射至同一潜在空间,支持双向查询。其方法融合自监督轨迹编码器、冻结的大规模语音编码器、对比联合嵌入及基于归一化流的双射升维。实验在旧金山湾区展开,该区域集中了主要机场,兼具商业与通用航空的低空交通。最后,本文发布一个新型语音-轨迹配对数据集,并报告了跨模态检索、消融实验与潜空间分析结果。
原文摘要 · Abstract (English)
As air traffic volumes in the National Airspace System continue to expand, in particular in the low altitude airspaces, the need for scalable decision support tools used by air traffic controllers will also require more development. This article introduces Voice-to-Trajectory for Air Traffic Control, a joint voice communication-flight trajectory data embedding framework, that can be a component of situational awareness in congested airspaces, and assist the development of tools for ATC as they reason in real-time over Automatic Dependent Surveillance-Broadcast trajectories, or the intent expressed by pilots in natural language. We show that these data modalities are not independent and represent a common physical referent: an aircraft flying through the airspace. V2TATC maps a voice instruction and the trajectory of the addressed aircraft to nearby points in a single latent space that can be queried in both directions. It combines a self-supervised trajectory encoder, a frozen large-scale speech encoder, a contrastive joint embedding, and a bijective lifting via normalizing flows. We demonstrate V2TATC's effectiveness on the San Francisco Bay Area, for its concentration of major airports, and its mix of commercial and general aviation low altitude traffic. Lastly, we release a novel paired voice-trajectory dataset, and report experiments on cross-modal retrieval, ablations, and latent-space analysis.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。