用合成数据提升机器人抓取的现实适应性,成功率超80%。
HyperSim: A Holistic Sim-To-Real Framework For Robust Robotic Manipulation

- 三重架构:高保真仿真、对抗轨迹生成、虚实联合训练
- 在真实环境中完成400次任务,成功率达80%~95%
- 对抗训练使动态干扰下的完成率提升35%
扩大数据量与多样性是实现具身智能泛化的关键。尽管合成数据生成可替代昂贵的物理数据采集,但将机器人抓取策略从仿真迁移到真实世界(模拟到现实)仍面临显著的域差异挑战。本文提出HyperSim,一个覆盖合成数据生成、策略训练到无缝部署的完整框架。通过三个核心支柱系统性弥合模拟到现实的差距:高保真环境构建、对抗性轨迹生成和虚实联合训练。这些模块共同通过提升视觉保真度、扩展数据覆盖范围、强制域不变表征来缓解域偏差。我们在两个代表性抓取模型上进行大规模实验,共执行400次真实任务。基于三项细粒度指标评估,完整流程在ACT和π₀模型上分别达到80%和95%的模拟到现实成功率。此外,在对抗性轨迹上训练的策略对动态不确定性表现出更强鲁棒性,物理扰动下完成率高出35%。
原文摘要 · Abstract (English)
Scaling data volume and diversity is critical for generalizing embodied intelligence. While synthetic data generation offers a scalable alternative to expensive physical data acquisition, transferring robotic manipulation policies from simulation to the real world (sim-to-real) remains a formidable challenge due to the domain gap. This paper presents HyperSim, a holistic framework spanning from synthetic data generation to policy training and seamless real-world deployment. To systematically bridge the sim-to-real gap, HyperSim is realized through three core pillars: high-fidelity environment synthesis, adversarial trajectory generation, and sim-and-real co-training. Collectively, these modules address domain discrepancies by enhancing visual fidelity, expanding data coverage, and enforcing domain-invariant representations. We rigorously validate HyperSim through a large-scale empirical study involving 400 real-world task executions across two representative manipulation models. Assessed across three fine-grained metrics, our complete pipeline achieves remarkable sim-to-real success rates of 80% and 95% with ACT and π_{0}, respectively. Furthermore, policies trained on our adversarial trajectories exhibit significantly enhanced robustness against dynamic uncertainties, achieving a 35% higher completion rate under physical perturbations.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。