Flow-based Policy Adaptation without Policy Updates
TL;DR: GLOVES uses flow matching to selectively pull out-of-distribution actions from humans or pretrained policies toward expert behavior—without updating the underlying policy. With limited demonstrations, it improves imperfect imitation-learning and VLA agents across simulated and real-robot tasks.