
Reinforcement Learning for Real-Time Vision-Language-Action Policies
arXiv preprint arXiv:2609.18207, 2026
Presents a framework for reinforcement learning fine-tuning of real-time policies through decoupled slow action generation and fast, reactive editing, combining the reliability of RL with the reactiveness required for dynamic control.










