用物理滤波反馈提升机器人在少数据下的泛化能力
Physics Filtering Favors the Generalization of Robot Learning

- 引入物理滤波模块PhyFilter,通过物理约束修正学习误差
- 仅用少量训练数据,使四足机器人适应未知地形与负载变化
- 适合需要强泛化能力的机器人系统,尤其数据获取困难场景
生物体通过内在物理结构和持续反馈学习,在未知环境中表现出卓越适应性。赋予机器人类似泛化能力对真实世界可靠运行至关重要。尽管现有方法依赖大规模训练数据提升泛化性能,但机器人领域难以实现如大语言模型般的数据规模,因真实世界示范采集成本高、速度慢。本文提出相反策略:即使数据有限,仅通过物理滤波反馈机制PhyFilter,即可有效应对动态不确定性。PhyFilter作为轻量级、模型无关模块,其参数可通过自动学习算法优化,无需人工调参,可无缝集成至多种机器人策略中。我们在四个代表性机器人系统上验证:四足机器人能泛化至未知地形、载荷及速度范围;无人机可在未见风扰下飞行;空中机械臂在风力与质量不确定下仍可实现厘米级空中抓取;加速度差分器在分布偏移下保持鲁棒性。结果表明,物理滤波反馈可成为替代海量数据扩展的强大方案。
原文摘要 · Abstract (English)
Living organisms exhibit extraordinary adaptability to unseen environments through their intrinsic physical structures and lifelong feedback-driven learning. Endowing robots with comparable generalization is critical for reliable operation in the real world. While recent approaches attempt to improve generalization by scaling training data, such strategies remain impractical for robotics, where collecting real-world demonstrations at the scale of large language models is prohibitively costly and slow. Contrary to this reliance on massive datasets, we show that robots can generalize effectively under dynamics uncertainties even with limited training data by leveraging a feedback mechanism, namely PhyFilter, that corrects learning outputs with physics-filtered learning residuals. PhyFilter operates as a lightweight, model-agnostic module whose parameters can be automatically optimized through an auto-learning algorithm, eliminating manual tuning and enabling seamless integration with diverse robot policies. We validate PhyFilter across four representative robotic systems, demonstrating that it enables quadruped robots to generalize to unseen terrains, payload variations, and speed ranges; drones to flight under unseen wind disturbances; aerial manipulators to achieve centimeter-level in-air capture despite wind and mass uncertainties; and acceleration differentiators to remain robust with distribution shift. These results show that physics-filtered feedback can serve as a powerful alternative to massive data scaling.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。