AuraLuxMuse: Adaptive Fusion Modeling for Aesthetic Stage Lighting Design with Music and Expert Guidance
Junyu Deng, Jiale Cao, Mengtian Li, Zhongxia Ji, Ruhua Chen, β¦
https://t.co/wUo65JJ3qC [ππ.πΌπΌ ππ.π°πΈ]
π¬Accepted to appear in SIGGRAPH Asia 2026 Conference Papers
MemoCare: An Interactive Multimodal Mobile System for Automated Cognitive Screening
Duy-Cat Can, Mau Minh Phuc Le, Tuan-Khoa Hoang, Hai-Dang Nguyen, Trung-Hieu Do, Dang Minh Ly, β¦
https://t.co/cDMGC0XrKN [ππ.πΌπΌ ππ.π²π ππ.π·π²]
π¬Submitted to the MMM 2027 Demo Track
From Expression to Reaction: Role-aware Visual Transfer and Stimulus-guided Reasoning for Interlocutor Emotion Recognition
Wei Wang, Zhaowu Li, Jianjie Luo, Fu Lee Wang, Lap-Kei Lee, Zhenguo Yang
https://t.co/01HGuUMn6i [ππ.πΌπΌ ππ.π²π ]
Toward Generative Video Communication: A Dual-Stream Digital Transmission Framework
Bingyan Xie, Longyu Zhou, Tianhao Liang, Yongpeng Wu, Zehui Xiong, Wenjun Zhang, Tony Q. S. Quek
https://t.co/8qnRsmMC9n [ππ.πΌπΌ]
π¬Accepted by the IEEE Wireless Communications Magazine
Small Cues, Big Consequences: Learning Pivotal Cues for Multimodal Meme Classification
Akshit Sharma, Prashant W. Patil
https://t.co/euYUsdbxp0 [ππ.πΌπΌ ππ.π²π» ππ.π»πΆ]
π¬Accepted to EMNLP 2026 Findings
ROAM-ASD: Robust Open-World Active Speaker Detection with Flexible Multimodal Fusion
Pu Wang, Yujun Wang, Hugo Van hamme
https://t.co/Qxjg0CCX0o [ππ.πΌπΌ ππ.π²π ππ.ππ³ ππππ.π°π ππππ.πΈπ ]
π¬Submitted to IEEE ICASSP 2027
CogenPVG: Cognitive-Enhanced Reflective Multi-Agent Framework for Persuasive Video Generation
Yuntian Xiao, Shoulong Zhang, Wenfeng Song, Yan Wang, Yi Chen, Shuai Li
https://t.co/iE7ExlbDZI [ππ.πΌπΌ ππ.π°πΈ]
If You Hear It, Help Find It: Orthogonal Knowledge Distillation for Open-Vocabulary Audio-Visual Event Localization
Yi Xu, Cheng Chen, Wenzhuo Lei
https://t.co/GLJ1Mr5WrZ [ππ.πΌπΌ ππ.π²π ππ.π»πΆ ππ.ππ³]
π¬Accepted to ACM Multimedia 2026 (poster)