Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Incremental Learning of Retrievable Skills For Efficient Continual Task Adaptation

About

Continual Imitation Learning (CiL) involves extracting and accumulating task knowledge from demonstrations across multiple stages and tasks to achieve a multi-task policy. With recent advancements in foundation models, there has been a growing interest in adapter-based CiL approaches, where adapters are established parameter-efficiently for tasks newly demonstrated. While these approaches isolate parameters for specific tasks and tend to mitigate catastrophic forgetting, they limit knowledge sharing among different demonstrations. We introduce IsCiL, an adapter-based CiL framework that addresses this limitation of knowledge sharing by incrementally learning shareable skills from different demonstrations, thus enabling sample-efficient task adaptation using the skills particularly in non-stationary CiL environments. In IsCiL, demonstrations are mapped into the state embedding space, where proper skills can be retrieved upon input states through prototype-based memory. These retrievable skills are incrementally learned on their corresponding adapters. Our CiL experiments with complex tasks in Franka-Kitchen and Meta-World demonstrate robust performance of IsCiL in both task adaptation and sample-efficiency. We also show a simple extension of IsCiL for task unlearning scenarios.

Daehee Lee, Minjong Yoo, Woo Kyung Kim, Wonje Choi, Honguk Woo• 2024

Related benchmarks

TaskDatasetResultRank
Continual Imitation LearningLIBERO goal 1.0
FWT (Forward Transfer)66.8
34
Continual LearningLIBERO Object
FWT57
25
Continual Imitation LearningLIBERO long 1.0
Forward Transfer (FWT)65.8
23
Lifelong Imitation LearningLIBERO Goal
Forward Transfer (FWT)70.4
16
Continual LearningLIBERO Long
Forward Transfer (FWT)44
15
Continual LearningLIBERO Spatial
FWT38
14
Lifelong LearningLIBERO Goal
Forward Transfer (FWT)54
10
Lifelong Imitation LearningLIBERO Object
Forward Transfer (FWT)71.7
9
Task AdaptationEvolving World-Complete (Unseen)
FWT64.3
8
Lifelong Imitation LearningLIBERO-50
Forward Task Transfer (FWT)47.8
8
Showing 10 of 16 rows

Other info

Follow for update