Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics

About

Synthesizing realistic Human-Object Interactions (HOI) is critical for creating embodied avatars and functional virtual environments. However, current data-driven approaches primarily rely on motion capture datasets, which are expensive to scale and limited in functional diversity. Models trained with these datasets fail to generalize to unseen objects and maintain physical consistency over long horizons. In this paper, we propose a novel framework that leverages a physics simulator to overcome the data-scarcity bottleneck in HOI generation. Specifically, we propose a scalable pipeline, called \ours, which leverages policies trained with reinforcement learning in a physics simulator for task-oriented data generation and trains a generative model on the augmented dataset for generalizable HOI generation. To seamlessly utilize the synthetic data, we introduce a coarse-to-fine retargeting process that bridges the representation gap between the simplified model used in physics simulator and the standard parametric body models required for generative training. Validated through comprehensive experiments, our method demonstrates enhanced generalization to unseen objects and the capability of long-horizon generation, while exhibiting greater dynamic diversity and physical plausibility.

Shujia Li, Jianshu Hu, Haiyu Zhang, Yunpeng Jiang, Haoyuan Jin, Xinyuan Chen, Yaohui Wang, Yutong Ban• 2026

Related benchmarks

TaskDatasetResultRank
Human-Object Interaction SynthesisOMOMO Seen Objects (test)
Contact Recall (Crec)67
7
Human-Object Interaction SynthesisOMOMO (5 strictly unseen objects)
Recall (C)64
6
Showing 2 of 2 rows

Other info

Follow for update