How does a robot learn to recognize an object it’s never encountered? In this case, a demo can be worth a thousand words.
Q2RL: Reinforcement Learning to Improve Beyond Behavior Cloning
Presented at RSS 2026, Q2RL (Q-Estimation and Q-Gating from Behavior Cloning for Reinforcement Learning) is a method that avoids catastrophic forgetting when transitioning from offline...