Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Humanoid-DART: Humanoid Loco-Manipulation using Diffusion-guided Augmentation through Relabeling and Tracking

About

Imitating human demonstrations has emerged as a dominant paradigm for learning humanoid loco-manipulation policies. However, scaling these approaches remains challenging due to the high cost of collecting diverse demonstrations and the need for continual human intervention to correct policy failures. In this paper, we present a self-supervised framework that bootstraps from sparse demonstrations and progressively expands its behavioral repertoire, enabling the learning of a goal-conditioned policy that automatically explores the goal space with minimal expert supervision. Our approach combines diffusion-based trajectory generation with reinforcement learning, where the latter is used to track goal-conditioned trajectories produced by the diffusion model for a range of loco-manipulation skills. Through extensive ablation studies and comparisons with state-of-the-art methods, we demonstrate the effectiveness of our framework on multiple humanoid loco-manipulation skills.

Pranav Debbad, Kanish Thiagarajan, Victor Dh\'edin, Shafeef Omar, Majid Khadiv• 2026

Related benchmarks

TaskDatasetResultRank
Hand-offLoco-manipulation Task Suite 20 novel goals
Coverage51.5
3
KickLoco-manipulation Task Suite 20 novel goals
Coverage54.8
3
Pick-&-PlaceLoco-manipulation Task Suite 20 novel goals
Coverage96.4
3
PushLoco-manipulation Task Suite 20 novel goals
Coverage (%)61.1
3
Showing 4 of 4 rows

Other info

Follow for update