Announcing RoboFuME๐ค๐จ, a system for autonomous & efficient real-world robot learning that
1. pre-trains a VLM reward model and a multi-task RL policy from diverse off-the-shelf demo data;
2. runs RL fine-tuning online with the VLM reward model.
๐ https://t.co/Q3c0Em6xva
๐งตโ