Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipulation

About

Diffusion policies have achieved remarkable success in robotic manipulation, yet they often fail to satisfy strict physical constraints required for safe deployment. Existing approaches impose safety either prematurely during training or reactively via external guardrails at test time, limiting policy expressivity and overall scalability. We propose Physical safety Alignment for Constrained Trajectories (PACT), a self-evolving post-training framework that projects pretrained diffusion policies onto constraint-feasible regions without accessing demonstration data or task rewards. PACT distills constraint gradients into the diffusion model through a reverse-KL objective with dense supervision across timesteps. It incorporates a curriculum that progressively tightens constraints while maintaining theoretically bounded policy shift and monotone improvement, mitigating the safety-performance trade-off from catastrophic forgetting. On simulated and real-world embodied manipulation benchmarks, PACT significantly reduces safety violations by 31.0% on average while improving task success by 30.7%.

Lingxuan Wu, Zijian Zhu, Lizhong Wang, Chengyang Ying, Huayu Chen, Xiao Yang, Fangming Liu, Jun Zhu• 2026

Related benchmarks

TaskDatasetResultRank
Handover AppleRoboTwin
Success Rate86
8
handover blockRoboTwin
Success Rate0.75
8
Pick Diverse BottleRoboTwin
Success Rate76
8
Pick Dual BottleRoboTwin
Success Rate96
8
Place Dual ShoesRoboTwin
Success Rate70
8
Pour WaterRoboTwin
Success Rate86
8
stack blocksRoboTwin
Success Rate92
8
Showing 7 of 7 rows

Other info

Follow for update