Launching today: Layerize any image
Drag in an image and Moda will rebuild layers that you can edit directly.
Early users have been saying text reproduction is way better than Canva or Photoshop.
See it in action:
Introducing the Moda API: the design layer for agents, apps, and workflows.
Embed presentations, reports, and design creation right into your product or agent.
See what people are building with it 👇
Happy to announce TWIN was accepted to #CVPR2026!
We introduce a large-scale VQA dataset for fine-grained understanding in VLMs and show that post-training on this data improves fine-grained visual reasoning for OOD domains!
See everyone in Denver! 👋🏔️
(1/6) Do these images show the same vacuum cleaner? They are certainly similar, but a human will notice the differences in dustbin geometry, design, and color accents. In contrast, open-source VLMs struggle at this task. Our recent work TWIN poses the question: Can we fix this?
(6/6) This work was done in collaboration with two amazing undergrads @Amehta_2005, @rlin232, and my advisor @georgiagkioxari !
Project page: https://t.co/gqkqswlK4U
Paper: https://t.co/MgI17D14YA