arXiv:2603.25275cs.CV2026-03中稿 · CVPR被引 2

首个车机协同感知大规模真实数据集,提升复杂场景下目标检测能力

V2U4Real: A Real-world Large-scale Dataset for Vehicle-to-UAV Cooperative Perception

  • 车与无人机协同采集多模态数据,覆盖城市、校园、乡村多种场景
  • 含56,000帧激光雷达数据和70万标注3D边界框,支持长距与遮挡感知
  • 适用于自动驾驶、多智能体感知研究,推动跨视角协作感知发展

现代自动驾驶感知系统常受遮挡、盲区和传感范围限制。现有车车(V2V)与车路(V2I)协同感知虽有效缓解问题,但局限于地面协作,难以应对复杂环境中的大范围遮挡与远距离感知。为推进跨视角协同感知研究,我们提出V2U4Real,首个面向车机(V2U)协同感知的大规模真实世界多模态数据集。该数据集由搭载多视角激光雷达与RGB相机的车载平台与无人机共同采集,涵盖城市街道、大学校园及乡村道路,在多样交通场景下共包含超过56,000帧激光雷达数据、56,000组多视角图像以及700,000个标注的3D边界框,涉及四类物体。为支持多样化研究任务,我们建立了单智能体3D目标检测、协同3D目标检测与目标跟踪的基准评测体系。对多个前沿模型的全面评估表明,车机协同显著提升了感知鲁棒性与远距离感知能力。V2U4Real数据集及代码已开源:https://github.com/VjiaLi/V2U4Real。

原文摘要 · Abstract (English)

Modern autonomous vehicle perception systems are often constrained by occlusions, blind spots, and limited sensing range. While existing cooperative perception paradigms, such as Vehicle-to-Vehicle (V2V) and Vehicle-to-Infrastructure (V2I), have demonstrated their effectiveness in mitigating these challenges, they remain limited to ground-level collaboration and cannot fully address large-scale occlusions or long-range perception in complex environments. To advance research in cross-view cooperative perception, we present V2U4Real, the first large-scale real-world multi-modal dataset for Vehicle-to-UAV (V2U) cooperative object perception. V2U4Real is collected by a ground vehicle and a UAV equipped with multi-view LiDARs and RGB cameras. The dataset covers urban streets, university campuses, and rural roads under diverse traffic scenarios, comprising over 56K LiDAR frames, 56K multi-view camera images, and 700K annotated 3D bounding boxes across four classes. To support a wide range of research tasks, we establish benchmarks for single-agent 3D object detection, cooperative 3D object detection, and object tracking. Comprehensive evaluations of several state-of-the-art models demonstrate the effectiveness of V2U cooperation in enhancing perception robustness and long-range awareness. The V2U4Real dataset and codebase is available at https://github.com/VjiaLi/V2U4Real.

车机协同3D检测多模态数据集自动驾驶

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。