arXiv · 2409.08233
Towards Online Safety Corrections for Robotic Manipulation Policies
Abstract
Recent successes in applying reinforcement learning (RL) for robotics has shown it is a viable approach for constructing robotic controllers. However, RL controllers can produce many collisions in environments where new obstacles appear during execution. This poses a problem in safety-critical settings. We present a hybrid approach, called iKinQP-RL, that uses an Inverse Kinematics Quadratic Programming (iKinQP) controller to correct actions proposed by an RL policy at runtime. This ensures safe execution in the presence of new obstacles not present during training. Preliminary experiments illustrate our iKinQP-RL framework completely eliminates collisions with new obstacles while maintaining a high task success rate.
Explore related subjects
Keep this discovery
Ariana Spalter, Mark Roberts, Laura M. Hiatt. 2024-09-12. Towards Online Safety Corrections for Robotic Manipulation Policies. https://arxiv.org/abs/2409.08233
Cite the original work for its findings. Save a collection to share your selection of sources.