Meta发布500小时3D动作数据集,支持多人交互与情感对话研究。
Embody 3D: A Large-scale Multimodal Motion and Behavior Dataset
- 采集439人500小时多视角3D动作数据,超5400万帧
- 包含单人动作与多人互动、情绪对话等丰富场景
- 含手部追踪、身体形态、文本标注和独立音频流
Meta旗下Codec Avatars实验室推出Embody 3D,一个大规模多模态动作与行为数据集。该数据集涵盖439名参与者的500小时3D运动数据,通过多摄像机采集,共生成超过5400万帧的3D动作追踪数据。内容包括单人动作(如指令动作、手势、行走)及多人行为(如讨论、情绪化对话、协作活动、公寓式共居场景)。数据提供人体动作追踪(含手部与体态)、文本注释及每位参与者独立的音频轨道,支持多模态交互研究。
原文摘要 · Abstract (English)
The Codec Avatars Lab at Meta introduces Embody 3D, a multimodal dataset of 500 individual hours of 3D motion data from 439 participants collected in a multi-camera collection stage, amounting to over 54 million frames of tracked 3D motion. The dataset features a wide range of single-person motion data, including prompted motions, hand gestures, and locomotion; as well as multi-person behavioral and conversational data like discussions, conversations in different emotional states, collaborative activities, and co-living scenarios in an apartment-like space. We provide tracked human motion including hand tracking and body shape, text annotations, and a separate audio track for each participant.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。