Meet Blue Jay, our next-generation robotic system that's like having an extra set of hands for our teams 🤖. It picks, stows, and consolidates in one streamlined workspace—moving three separate stations into one.
This is @Figure_Robot III announced today. I’m sorry but every other form factor besides humanoids will be crushed 🦾🤖🦾
I keep hearing people starting specialized robotics companies, “like robots for laundry”. And the pitch is: “you don’t need a humanoid form factor to do laundry, it can be much simpler”. That’s true. But, humanoids can do everything and will become the default form factor and so will be mass manufactured and so will be extremely cheap compared to lower volume robots. Its so obvious, it will be humanoids doing laundry, not a specialized laundry robot 🤦♂️
The one main exception to that might be line work in a factory where you do just need the arms and you can mass manufacture them (a al @proceptionAI)
Missed the activation, but the set looked like a modern art installation.
Great job by Apple TV+. Good to know that The Big Idea is still alive and well in some marketing boardrooms.
🎥 Today we’re premiering Meta Movie Gen: the most advanced media foundation models to-date.
Developed by AI research teams at Meta, Movie Gen delivers state-of-the-art results across a range of capabilities. We’re excited for the potential of this line of research to usher in entirely new possibilities for casual creators and creative professionals alike.
More details and examples of what Movie Gen can do ➡️ https://t.co/M19x2ndwnr
🛠️ Movie Gen models and capabilities
Movie Gen Video: 30B parameter transformer model that can generate high-quality and high-definition images and videos from a single text prompt.
Movie Gen Audio: A 13B parameter transformer model that can take a video input along with optional text prompts for controllability to generate high-fidelity audio synced to the video. It can generate ambient sound, instrumental background music and foley sound — delivering state-of-the-art results in audio quality, video-to-audio alignment and text-to-audio alignment.
Precise video editing: Using a generated or existing video and accompanying text instructions as an input it can perform localized edits such as adding, removing or replacing elements — or global changes like background or style changes.
Personalized videos: Using an image of a person and a text prompt, the model can generate a video with state-of-the-art results on character preservation and natural movement in video.
We’re continuing to work closely with creative professionals from across the field to integrate their feedback as we work towards a potential release. We look forward to sharing more on this work and the creative possibilities it will enable in the future.