Introducing Waddle Labs: Claude Code for robots.
Connect our API to your robot and enter a prompt, then our agents write code to achieve the task in 20 minutes.
@yiding_song@theWaddleLabs
Introducing Waddle Labs: Claude Code for robots.
Connect our API to your robot and enter a prompt, then our agents write code to achieve the task in 20 minutes.
@yiding_song@theWaddleLabs
Sometimes code-as-policy works best (eg. repetitive tasks in factories), sometimes robot models work best (eg. walking a humanoid, controlling 5-fingered hand).
People over-index on models, but we still need em. What's surprising is that code-as-policy and LLM controlling robots is far more capable than expected. Ths should unlock faster deployments and automated data collection.
@hatebunnyplzzz@yiding_song@theWaddleLabs Cap-X is a foundational paper for code-as-policy! We didn't know "code-as-policy" was a thing in research when we started. But once we read about Cap-X it helped us design some of our primitives (see https://t.co/KEmKdT8Bai)
@tomcocobrico@NickADobos Nah thanks Jeffrey you raise a very valid point (also its X).
We've been using code-as-policy to collect tons of data internally - let's see where this goes
@NickADobos What's surprising is that code-as-policy and LLM controlling robots is far more capable than expected.
I expect this to unlock faster deployments and automated data collection.
@NickADobos I don't think so.
There’s no one path to robot intelligence. Sometimes code-as-policy works best (eg. repetitive tasks in factories), sometimes robot models work best (eg. walking a humanoid, controlling 5-fingered hand).
People over-index on models, but we still need em.
@Sujeetsoni123@yiding_song@theWaddleLabs LLMs definitely suck as physics. Maybe the root cause is that they suck at perception. We tried many models (every gemini, Nvidia Cosmos, you name it) and they all hallucinate an unbearable amount on videos.
@BrutalCaeser@yiding_song@theWaddleLabs In all the demos in the video there were no robot models (eg. ACT, VLAs). This goes to show that code-as-policy works in diverse situations.
I believe that robot models will be crucial for dexterous tasks, it's just that they're not the only way towards physical intelligence.
@Sujeetsoni123@yiding_song@theWaddleLabs I thought about duplex models as a way to improve accuracy actually. Because turn-based models miss changes in the environment unless they call perception tools intentionally. So full-duplex would make it easier to catch things like wind blowing over a water bottle.
@Sujeetsoni123@yiding_song@theWaddleLabs A big help would be better full duplex vision models (like what Thinking Machines or OAI have but for perception in particular) - lmk if you have ideas on this!
@Evose_AI@yiding_song@theWaddleLabs Good question - a lotta our work has gone into this. One thing we are working on: can we use sim / world models to evaluate programs before they run irl?
@Ken_Goldberg Many robot companies can make cool demos but we need standardised benchmarks to tell which robots are useful for society vs which just look cool
@xennygrimmato_@Figure_robot To be fair figure’s robots look hella sci fi. If they did a demo with an army of humanoids marching we wouldn’t be able to replicate that