VITAL: Interactive Few-Shot Imitation Learning via Visual Human-in-the-Loop Corrections

📅 2024-07-30
📈 Citations: 1
✨ Influential: 0
📄 PDF
🤖 AI Summary
To address the high cost and limited scale of high-quality real-world demonstration data in robotic imitation learning, this paper proposes a closed-loop learning framework integrating few-shot learning, simulation augmentation, and human-in-the-loop error correction. The method leverages a small set of real demonstrations (3–5 trials) augmented with synthetic data generated via GANs or diffusion models, followed by online policy fine-tuning to enable zero-shot cross-task transfer. Its key contributions are: (1) a novel vision-guided real-time human correction mechanism supporting natural multimodal interaction; (2) an end-to-end control pipeline built on ROS and PyTorch; and (3) empirical validation across manipulation tasks—including bottle collection, stacking, and hammering—achieving >92% success rates. Crucially, the framework transfers zero-shot to an unseen task—beverage tray arrangement—with 86% success, significantly outperforming pure-simulation baselines.

Technology Category

Intelligent Robots: ManipulationMachine Learning: Imitation Learning & Inverse Reinforcement LearningHumans and AI: Human-Aware Planning and Behavior Prediction

Application Category

Responsible Web: Machine-in-the-loop, human agency and autonomyEconomics, Online Markets and Human Computation: Humans versus LLMs for data annotation and labelingSearch and Retrieval-Augmented AI: Web learning to rank, online learning, and counterfactual learning for ranking
📝 Abstract
Imitation Learning (IL) has emerged as a powerful approach in robotics, allowing robots to acquire new skills by mimicking human actions. Despite its potential, the data collection process for IL remains a significant challenge due to the logistical difficulties and high costs associated with obtaining high-quality demonstrations. To address these issues, we propose a large-scale data generation from a handful of demonstrations through data augmentation in simulation. Our approach leverages affordable hardware and visual processing techniques to collect demonstrations, which are then augmented to create extensive training datasets for imitation learning. By utilizing both real and simulated environments, along with human-in-the-loop corrections, we enhance the generalizability and robustness of the learned policies. We evaluated our method through several rounds of experiments in both simulated and real-robot settings, focusing on tasks of varying complexity, including bottle collecting, stacking objects, and hammering. Our experimental results validate the effectiveness of our approach in learning robust robot policies from simulated data, significantly improved by human-in-the-loop corrections and real-world data integration. Additionally, we demonstrate the framework's capability to generalize to new tasks, such as setting a drink tray, showcasing its adaptability and potential for handling a wide range of real-world manipulation tasks. A video of the experiments can be found at: https://youtu.be/YeVAMRqRe64?si=R179xDlEGc7nPu8i
Problem

Research questions and friction points this paper is trying to address.

Reducing high costs of data collection for imitation learning
Enhancing robot policy robustness with human corrections
Generalizing learned skills to new real-world tasks
Innovation

Methods, ideas, or system contributions that make the work stand out.

Large-scale data generation via simulation augmentation
Visual processing with affordable hardware for demonstrations
Human-in-the-loop corrections enhance policy robustness
🔎 Similar Papers
2024-04-27Autonomous RobotsCitations: 2
💼 Related Jobs
No related jobs found.
University of Groningen | University of Edinburgh
H
H. Kasaei
Department of Artificial Intelligence, University of Groningen, The Netherlands
M
M. Kasaei
School of Informatics, University of Edinburgh, UK