i gave astra a robot, a paint brush, and a camera then asked it to paint the golden gate bridge in real life!
it figured out how to control the robot, and progressively got better throughout its attempts. the timelapse is sick
Action chunking — especially executing long action sequences open-loop — is widely used in imitation learning for robotic manipulation. Why is it so effective and do we really need it? We find a key reason:
Long open-loop execution helps short-context policies imitate non-Markovian experts.
With this insight, we show how to move beyond open-loop execution: extending policy context restores reactivity while achieving even higher task performance.
🧵(1/5)
Gradient descent extends easily to Riemannian manifolds. Does Nesterov's accelerated gradient method generalize to Riemannian manifolds? The answer turns out to be rather intricate. First set of answers in our colt18 paper: https://t.co/bXZNnA19Ds