Check our latest updates and improved model for PAIR Diffusion: A Comprehensive Multimodal Object-Level Image Editor🚀🚀
Project page: https://t.co/6aIxv2MAy2
ArXiv: https://t.co/ShNOGF7Ntz
We show that 👇👇
PAIR-Diffusion: Object-Level Image Editing with Structure-and-Appearance Paired Diffusion Models @Gradio demo is out on @huggingface Spaces
demo: https://t.co/hR63tPMl5i
We're making it easier for @GeminiApp to work across Google. Three weeks ago, it was Google's Shopping Graph and the 50 billion product listings there.
Today, it's Gemini 🤝 Google Maps!
With generative video models rising, people ask: do we still need 3D/4D? If you want cheap, realtime, interactive experiences — the answer is yes.
We just launched Animated Selfie Attachments in Lens Studio 5.17: generate interactive on-device experiences from a text prompt.
Very happy to bring generative 4D to the hands of our users!
This was a great collab with Snap Research and would not have been possible without @wang12_gordon and @ashmrz10’s paper that was published this year on NeurIPS!
https://t.co/RlGwmzNGea
With generative video models rising, people ask: do we still need 3D/4D?
If you want cheap, realtime, interactive experiences — the answer is yes.
We just launched Animated Selfie Attachments in Lens Studio 5.17: generate interactive on-device experiences from a text prompt.
Possibly, we can only store high level information like semantics and also some 3D representation in efficient manner and use generative models to decode them to high quality 3D representation of the world where our system can take actions.
Interesting read. Further, I recently noticed that though the brain only takes 20W of energy, it can have 1-2 petabytes of storage!
Moving forward I think we should relax some constraints on long term memory a world model can store.
Even with this large storage representation is highly important. If we record 30 fps video in 512 x 512 resolution we will only be able to record ~1-2 year of data even if we use whole of storage of brain 1-2 PB. Generative models might be useful for compressing data efficiently
This morning on the way to school, my 8-year-old daughter and I talked about fame and impact. We started with MrBeast and internet celebrities—whose work she knows well—but then I introduced Einstein, whose discoveries shaped the technology we use every day.
Her big question: ‘Do I have to change the world, Dad?’
My answer: you don’t have to aim for that. We just want you to be happy. Do the right thing, be kind, and follow what you love—the rest will follow naturally.
In today’s noisy world of rapid AI and technological breakthroughs and constant announcements, external pressures and expectations can make it feel like only big, visible achievements matter. But doing the right thing—and pursuing what you love, the things that keep you up at night and get you out of bed in the morning—even without instant recognition, is the only path to true happiness and serenity.
A reminder I share with my children, my students—and myself.
Hi all, I will be at CVPR in Nashville from 10-15 June. Lets meet!
Also drop by our paper Wonderland: Navigating 3D Scenes from a Single Image
https://t.co/Qe43iqeNrJ
When - Friday morning session
Where - ExHall D Poster #59
#CVPR25
talk by @jon_barron . Completely agree, further if we move towards spatial computing 3D would be definitely needed but again a long term bet.
Full talk: https://t.co/7tUzEKrA3K
Wondering what's happening with NATTEN in 2025?
Check out Generalized Neighborhood Attention!
Spoiler: NATTEN gets a new stride parameter, we made a simulator for all your analytical studies, AND a Blackwell kernel!
Keep reading for more...
(1 / 5)