用视觉注意力与多模态数据实现贴身衣物的自适应穿脱
Close-Fitting Dressing Assistance Based on State Estimation of Feet and Garments with Semantic-based Visual Attention
- 融合视觉、力觉与深度信息,实时估计足部与衣物状态
- 在10名真人上成功穿袜,成功率高于现有方法
- 适合老龄化社会中需个性化护理的场景
随着人口老龄化加剧,护理人力短缺问题日益突出。贴身衣物如袜子的穿戴尤为困难,需精细调节力度以应对皮肤摩擦或勾挂,同时考虑衣物形状与位置。本文提出一种基于语义视觉注意力的多模态方法,结合机器人摄像头图像、关节角度、关节力矩及触觉力,实现对个体差异的自适应力交互。通过引入基于物体概念的语义信息,而非仅依赖RGB数据,提升了对未见足型与背景的泛化能力;深度数据则用于推断袜子与脚之间的相对空间关系。为验证语义建模能力与安全性,训练数据使用假人采集,后续在真人上进行实验。结果表明,该模型能准确估计足部与衣物状态,成功为10名参与者完成穿袜,成功率优于Action Chunking with Transformer和Diffusion Policy。
原文摘要 · Abstract (English)
As the population continues to age, a shortage of caregivers is expected in the future. Dressing assistance, in particular, is crucial for opportunities for social participation. Especially dressing close-fitting garments, such as socks, remains challenging due to the need for fine force adjustments to handle the friction or snagging against the skin, while considering the shape and position of the garment. This study introduces a method uses multi-modal information including not only robot's camera images, joint angles, joint torques, but also tactile forces for proper force interaction that can adapt to individual differences in humans. Furthermore, by introducing semantic information based on object concepts, rather than relying solely on RGB data, it can be generalized to unseen feet and background. In addition, incorporating depth data helps infer relative spatial relationship between the sock and the foot. To validate its capability for semantic object conceptualization and to ensure safety, training data were collected using a mannequin, and subsequent experiments were conducted with human subjects. In experiments, the robot successfully adapted to previously unseen human feet and was able to put socks on 10 participants, achieving a higher success rate than Action Chunking with Transformer and Diffusion Policy. These results demonstrate that the proposed model can estimate the state of both the garment and the foot, enabling precise dressing assistance for close-fitting garments.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。