⚡ One policy, millions of embodiments, over 200 robot models. Can we add yours?
We're building γ₀, a generalist RL policy for motion control trained across millions of randomized embodiments derived from a growing collection of more than 200 robot models.
Our paper on Streaming RL under Partial Observability won Best Paper Award at the Big Worlds Workshop at #RLC2026. We enable memory in streaming RL using recurrent trace units.
Project page: https://t.co/gXGExiEIIz
Collaborators: @__noahfarr__ *, @CarloDeramo , @Jan_R_Peters
$6.7T is going toward data center capex through 2030. It still takes 10-12 months to design data centers.
@MarengoYCS26 is the engineering firm for accelerated data center development. Site Due Diligence through FEED and Permitting Design in half the traditional time and cost.
Data center design takes 10-12 months, and because every concept is manually produced, design firms explore only 3-4 options per site, leading to suboptimal designs, risks that surface late, and delays that cost developers.
Marengo reduces those traditional pre-construction cycles down to 5-6 months by automating iterative engineering processes, allowing for faster design iterations, earlier constraint and risk visibility, and higher-confidence decisions, turning early development speed into a lasting advantage.
Our professional engineering team uses in-house-built AI tools to accelerate the work, validating and choosing the best designs for you. They present you with a decision-ready package built from data and expertise, at every stage: Due Diligence, Feasibility, Concept, and FEED/Permitting Design.
Before YC, @gadmarconi and I helped build and deploy on-site data center infrastructure, worked on the world's most powerful magnets (56MW) at the 1.4GW MagLab facility, spent time as an AI software engineer at a nuclear and energy EPC, as well as published multiple papers on AI for complex engineering design. We know firsthand the manual processes that slow down design and lead to suboptimal decisions.
Are you a data center developer? Send a site. Let us show you what accelerated engineering looks like on one of your projects. https://t.co/mPaIKV5KGX
This article on respiratory systems is really interesting. I had no idea that birds and crocodiles shared a breathing system that goes all the way back to the common ancestor of all archosaurs.
https://t.co/xBfUzui1Va
credit: @EBKaczmarek, J. Phillips, B. Ryerson
Real-world online Reinforcement Learning from Vision + ForceTorque + Proprioception --> With MSDP we achieve ~90% success rate in only 35min of online training!
Checkout the Preprint here: https://t.co/bmme0ESX0v
Thanks to @FatAndFurious42@GabrieleTiboni1@GeorgiaChal!
Totally stoked that people like our paper! If you want to know more about recursive reasoning in multi-agent RL, come to my poster at #EWRL2025 at 11:00 ♻️
In this paper, the authors compute the gradient update of the policy of one agent by accounting also for the update of all other agents. I feel this is a fairly general idea that could be applied to most multi-agent RL algorithms.
🔗https://t.co/nLLmmgpMIh
@f14bertolotti Thanks for sharing our paper, I'm glad to see that people find the idea of recursive reasoning interesting!
I'll be presenting this at the European Workshop on RL this week for anyone who's interested in knowing more.
What does BAMF mean?
In Germany it means 'Bundesamt für Migration und Flüchtlinge' (Federal Office for Migration and Refugees), but whenever I hear it, I can't help but imagine the sound Nightcrawler from X-Men makes when he teleports
#germany#xmen#comics
Very excited to present 🎉Eau De Q-Network🎉 on Thursday at RLDM Poster #28
🔍Eau De Q-Network gradually prunes the network weights at the agent's learning pace, ultimately reaching a final sparsity level that is discovered by the algorithm!🔎
👉📰 https://t.co/4MV6noQ7Hd
"...for there is no greater glory that can befall a man living than what he achieves by speed of his feet or strength of his hands".
The Odyssey, book 8 line 147 (Richmond Lattimore's English translation)
#homer#odyssey
How do you tune the hyperparameters of your RL agent? 🤔
Come and chat with me about it tomorrow afternoon at poster 397 of @iclr_conf !
I will be presenting⚡️ Adaptive Q-Network⚡️a method that adaptively selects the hyperparameters of your RL agent during training 🔥