Read how we enabled a robot to reliably wipe up crumbs and spills with an approach for robotics applications in complex environments that uses an #RL policy (trained with a stochastic differential equation simulator) followed by a trajectory optimizer. → https://t.co/Iw7pjVBjac