arXiv:2409.14393cs.AIcs.RO2024-09被引 161

用遮蔽动作补全统一控制虚拟角色,支持多种指令自由切换。

MaskedMimic: Unified Physics-Based Character Control Through Masked Motion Inpainting

论文配图:MaskedMimic: Unified Physics-Based Character Control Through Masked Motion Inpainting
图 1 · 摘自论文原文
  • 将物理角色控制转为遮蔽动作补全任务,统一多模态输入
  • 支持关键帧、文本、场景信息等混合输入,生成连贯动画
  • 无需繁琐奖励设计,适合交互式游戏与虚拟体验开发

构建一个能适应多种场景的通用物理角色控制器是角色动画的重要前沿。理想的控制器应支持稀疏目标关键帧、文本指令和场景信息等多种控制方式。现有方法多专注于特定任务,难以泛化。本文提出MaskedMimic,将物理角色控制建模为通用的动作补全问题。核心思想是训练单一统一模型,从部分遮蔽的运动描述(如缺失关键帧、物体状态、文本描述或其组合)中合成完整动作。通过利用运动捕捉数据并设计可扩展的训练方法,该模型学习到一个无需为每种行为单独设计奖励函数的物理控制器。结果表明,该控制器支持多种控制模态,并可在不同任务间无缝切换。通过统一的动作补全框架,MaskedMimic实现了可动态适应复杂场景、按需生成多样化动作的虚拟角色,显著提升交互性与沉浸感。

原文摘要 · Abstract (English)

Crafting a single, versatile physics-based controller that can breathe life into interactive characters across a wide spectrum of scenarios represents an exciting frontier in character animation. An ideal controller should support diverse control modalities, such as sparse target keyframes, text instructions, and scene information. While previous works have proposed physically simulated, scene-aware control models, these systems have predominantly focused on developing controllers that each specializes in a narrow set of tasks and control modalities. This work presents MaskedMimic, a novel approach that formulates physics-based character control as a general motion inpainting problem. Our key insight is to train a single unified model to synthesize motions from partial (masked) motion descriptions, such as masked keyframes, objects, text descriptions, or any combination thereof. This is achieved by leveraging motion tracking data and designing a scalable training method that can effectively utilize diverse motion descriptions to produce coherent animations. Through this process, our approach learns a physics-based controller that provides an intuitive control interface without requiring tedious reward engineering for all behaviors of interest. The resulting controller supports a wide range of control modalities and enables seamless transitions between disparate tasks. By unifying character control through motion inpainting, MaskedMimic creates versatile virtual characters. These characters can dynamically adapt to complex scenes and compose diverse motions on demand, enabling more interactive and immersive experiences.

角色控制动作补全物理模拟多模态

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。