arXiv · 2610.00542
Does Continual Imitation Learning Remain Grounded? A Language-Perturbed Benchmark for Robotic Task Retention
Abstract
Continual imitation learning evaluates whether a robot can learn new knowledge without forgetting previously learned skills. However, retaining task performance does not ensure the behavior remains grounded in language because policies may rely on scene cues, object associations, or memorized task structure. We introduce a benchmark protocol to study how language-guided behavior changes as robotic policies learn successive tasks. We construct meaning-preserving and meaning-changing instruction variants for the Goal, Spatial, Object, and Long suites of LIBERO. Policy experiments focus on LIBERO-Goal, evaluating Original and Paraphrase instructions after each continual-learning stage. We compare representative continual imitation learning methods under their original assumptions while separating task competence from language sensitivity. The proposed diagnostics complement standard learning and forgetting metrics by measuring semantic robustness, goal adaptation, and language sensitivity. Results show that strong continual-learning performance does not always translate to reliable language grounding, and our diagnostics help determine whether retained skills remain correctly guided by their instructions. Additional materials are available at https://sites.google.com/view/stillgrounded
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Siddeshwar Raghavan, Ziqin Yuan, Fengqing Zhu, Byung-Cheol Min. 2026-09-30. Does Continual Imitation Learning Remain Grounded? A Language-Perturbed Benchmark for Robotic Task Retention. https://arxiv.org/abs/2610.00542
Cite the original work for its findings. Save a collection to share your selection of sources.