🔥Vision as Unified Multimodal Generation🔥
🎯SenseNova-Vision🎯 unifies vision tasks (e.g., detection, keypoints, segmentation, depth, surface normals, point maps, and camera pose) as unified multimodal model (UMM) generation *with SOTA results*
- Code: https://t.co/a1WEmDLVe4
Thanks for reposting, Songyou! Vision Banana is great and really inspiring! Happy to take this direction one step further by bringing vision closer to foundation models. Looking forward to more discussions!
The full version of our eLife work, BEN, is now available. We design BEN as an open-source software with interfaces compatible with popular neuroimaging toolboxes, including FSL, AFNI, FreeSurfer, ANTs, and SPM. You are welcome to use it freely. https://t.co/ELzSZLqXUf
In collaboration with Dr. Tingying Peng, our work has just been published in IEEE Transactions on Medical Imaging. Hopefully, our proposed model, MouseGAN++, will benefit the scientific community of rodent neuroscience. Read the article: https://t.co/1a8an8P6FM