Visual thought shouldn’t be trapped in discrete text space!
Thrilled that CoVT (Chain-of-Visual-Thought) was selected for an oral presentation at #ECCV2026! 🎉
CoVT lets VLMs reason in continuous visual space, beyond discrete text tokens
Code and model: https://t.co/8DoQLzM7Lm
🥳 Excited to share that our paper, “Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens,” (CoVT) has been selected for an oral presentation at #ECCV2026!
CoVT is only a starting point: we are exploring more possibilities for visual reasoning, with the hope of further advancing breakthroughs in multimodal intelligence.
Project Page: https://t.co/iPxu8ipyLH
Code: https://t.co/d2oXBsROBH
Paper: https://t.co/lNGaO2fLlR
Huge thanks to @jiaxin_ge_ , @XDWang101 , @xkungfu , and all our amazing collaborators for their guidance and support!