Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

World Action Models Enable Continual Imitation Learning with Recurrent Generative Replays

About

Going beyond predicting robot actions, World Action Models (WAMs) can also generate future visual observations. We build on this generative capability to propose Recurrent Generative Replay (REGEN), a continual imitation learning framework that synthesizes pseudo-replay trajectories, enabling a robot policy to rehearse previously learned tasks without storing their original human demonstrations. During continual adaptation, REGEN recursively queries the WAM to synthesize pseudo-replay trajectories conditioned only on prior task instructions and current-task observations. Experiments in both simulation and real-world manipulation settings show that REGEN reduces catastrophic forgetting by up to $50\%$ relative to sequential fine-tuning, while approaching the performance of privileged experience replay methods that require access to real replay data. Finally, we analyze the factors limiting generated replay, identifying long-horizon visual degradation and action-observation inconsistency as the primary bottlenecks. Our results establish WAMs as a promising foundation for continual robot learning without stored demonstrations.

Manish Kumar Govind, Dominick Reilly, Smit Patel, Hieu Le, Srijan Das• 2026

Related benchmarks

TaskDatasetResultRank
Continual Imitation LearningLIBERO Spatial 13 (test)
Forward Transfer (FWT)87.2
7
Continual Imitation LearningLIBERO-Object 13 (test)
Forward Transfer (FWT)95.3
7
Continual Imitation LearningLIBERO-Goal 13 (test)
Forward Transfer (FWT)90.6
7
Continual Robot ManipulationReal-world Single-arm Manipulation Tasks T1-T3 (test)
Forward Transfer (FWT)80
2
Showing 4 of 4 rows

Other info

Follow for update