arXiv:2505.07266cs.RO2025-05ICRA被引 2

BETTY数据集覆盖高速竞速场景,助力全栈自动驾驶算法训练与评估。

BETTY Dataset: A Multi-modal Dataset for Full-Stack Autonomy

  • 采集6种竞速环境下的多模态数据,涵盖感知到控制全栈输出。
  • 包含13小时、32TB数据,最高时速达63米/秒,含失稳、打滑等极端状态。
  • 适合研究高动态场景下感知、预测与控制的联合优化,尤其适用于竞速类自动驾驶。

我们提出BETTY数据集,这是一个大规模多模态数据集,由多辆自动驾驶赛车在多种环境中收集,旨在支持监督与自监督的状态估计、动力学建模、运动预测、感知等任务。现有大规模自动驾驶数据集主要聚焦于感知、规划与运动预测等任务。本工作通过整合所有传感器输入、软件栈输出、语义元数据及真实标注信息,推动多模态数据驱动方法的发展。数据集涵盖4年积累,目前达13小时以上、32TB,覆盖6种多样化竞速环境,包括高速椭圆赛道、高纵向与横向加速度道路,以及无GPS信号的复杂场景。数据捕捉了高度动态状态,如63米/秒碰撞、轮胎打滑及极限稳定性运行。该数据集为全栈自动驾驶流程的训练与测试提供了丰富跨模态、高动态数据支持,推动各算法性能达到极限。当前数据集可通过https://pitt-mit-iac.github.io/betty-dataset/获取。

原文摘要 · Abstract (English)

We present the BETTY dataset, a large-scale, multi-modal dataset collected on several autonomous racing vehicles, targeting supervised and self-supervised state estimation, dynamics modeling, motion forecasting, perception, and more. Existing large-scale datasets, especially autonomous vehicle datasets, focus primarily on supervised perception, planning, and motion forecasting tasks. Our work enables multi-modal, data-driven methods by including all sensor inputs and the outputs from the software stack, along with semantic metadata and ground truth information. The dataset encompasses 4 years of data, currently comprising over 13 hours and 32TB, collected on autonomous racing vehicle platforms. This data spans 6 diverse racing environments, including high-speed oval courses, for single and multi-agent algorithm evaluation in feature-sparse scenarios, as well as high-speed road courses with high longitudinal and lateral accelerations and tight, GPS-denied environments. It captures highly dynamic states, such as 63 m/s crashes, loss of tire traction, and operation at the limit of stability. By offering a large breadth of cross-modal and dynamic data, the BETTY dataset enables the training and testing of full autonomy stack pipelines, pushing the performance of all algorithms to the limits. The current dataset is available at https://pitt-mit-iac.github.io/betty-dataset/.

自动驾驶多模态数据竞速场景全栈系统

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。