Adding image generation makes your LLM dumber (MoE/MoT collapse)๐. We fixed it!
๐ชจRosetta (HKUST ๐ค Tencent Hunyuan Foundation Model Team) eliminates gradient conflicts to keep your LLM smart.
โก๏ธStrictly ZERO extra VRAM overhead.
๐ปSingle-GPU pretraining reproducible in 90 mins!
https://t.co/lPOgDhzWVp
https://t.co/At0KBi98iR
See the collapse for yourself.
๐In just 100 training steps, MoE/MoT's ARC scores plummet from ~68 โก๏ธ ~30. Rosetta stays rock solid.
๐Try our 1-command demo (~90 mins on 1 or 8x H20):
https://t.co/wenkhfHrgb
Thrilled to share our paper, Symbiotic-MoE, has been accepted to ECCV 2026! ๐๐๐
โ How do we unify Multimodal Understanding and Generation in native foundation models without catastrophic forgetting?
๐ก We propose Symbiotic-MoE, a native multimodal Mixture-of-Experts architecture. By leveraging Shared Experts as a global semantic bridge, we achieve symbiotic co-evolution and mutual enhancement between text/vision understanding and image generationโwith zero extra parameter overhead! ๐
๐ Paper: https://t.co/ECDP178x3F
This work was developed during my internship at @TencentHunyuan. Deeply grateful to the amazing team and colleagues there for the inspiring discussions and support! See you in Malmรถ, Sweden this Sept.! ๐ธ๐ช๐ฅณ
#ECCV2026 #ComputerVision #DeepLearning #MoE #GenerativeAI #UnifiedModel #LMM #ImageGeneration
#3DV2025#GaussianAvatarEditor ๐ฅณ๐ฅณ๐ฅณ Thrilled to share our 3DV paper, an innovative framework for text-driven editing of animatable Gaussian head avatars that can be fully controlled in expression, pose, and viewpoint. Code has released! ๐๐๐
https://t.co/7uvsMfa9hq
#CVRP2024#NeRF2NeRF ๐ฅณ๐ฅณ๐ฅณThrilled to share our high score 554 CVPR paper "GenN2N: Generative NeRF2NeRF Translation"๐๐๐, a unified framework for various NeRF2NeRF tasks like NeRF editing, colorization, super-resolution, inpainting. Code coming soon! ๐https://t.co/PElkYKhJJn
Excited to share ๐๐๐ซ๐๐ง๐๐๐ซ-๐-๐๐ข๐๐๐จ, accepted by #SIGGRAPH_Asia 2023, for generating fantastic and coherent videos with #Stable_Diffusion and #Ebsynth. Code is released and try it at: https://t.co/JveqAWH6M6
We can also jointly optimize geometry and appearance. In the example below, we determine the albedo and roughness of a Disney BSDF (shown under new viewing/illumination conditions). (6/8)
Two job opportunities at the Swiss government in the field of robotics:
1.) Job with focus on countering mini drones: https://t.co/lqqCstESps
2.) Job with focus on unmanned ground vehicles (MedEvac, Logistics, Search & Rescue, ...): https://t.co/CHCwTPQxSW
#jobs#swissrobotics
@NVIDIAAI Toronto AI Lab (https://t.co/RpusM4YOdk) is looking for motivated interns (undergrad./MSc/PhD) to join our group! We are working on 3D/4D vision, simulation for robotics/AV, animation, generative modeling, content creation, and more. Apply here: https://t.co/ZQnINBB0tu