π₯³ Happy to announce that StreamGaze is accepted to #CVPR2026!
π We introduce the first benchmark that evaluates gaze-guided temporal reasoning (past, present, and future) and proactive understanding for streaming video understanding. We find that all MLLMs fall far below human performance, particularly in temporal continuity, gaze grounding, and proactive prediction.
π Huge thanks to my last year's AdobeResearch team: Subhojyoti Mukherjee, Branislav Kveton, Ryan A. Rossi, Viet Lai, David Seunghyun Yoon, Trung Bui, Franck Dernoncourt, and my advisor Mohit Bansal π
Key contributions to the captioning modelΒ SlimVLM-3B, which takes aΒ kidβs sketch and generates text to send to Firefly for image and 3D Generation. Thanks to amazing collaborator Viet and leader Trung.
π Thrilled to see the launch of Project Aqua! Itβs incredibly rewarding to see a product I contributed to make its way into the world. π
https://t.co/llCNtkDv1t
π Another presentation is happening today! Come stop by our poster session π
Program Synthesis via Test-Time Transduction
π Paper: https://t.co/T8HIViPSIn
π Thu, Dec 4, 2025
β° 4:30 PM β 7:30 PM PST
π Exhibit Hall C, D, E β #2005
#NeurIPS2025
Come to our poster session for a discussion! π
Offline RL by Reward-Weighted Fine-Tuning for Conversation Optimization
π Paper:: https://t.co/OSfCAKr13x
π Thu, Dec 4, 2025
β° 11:00 AM β 2:00 PM PST
π Exhibit Hall C, D, E β #210
#NeurIPS2025
Adobe researchers are sharing groundbreaking work this week at the premiere AI technology conference NeurIPS2025! We are proud to be presenting 24 papers that connect AI to the future of creativity, design, and data intelligence. We explore generative model innovation, efficient learning, personalized experiences, and multimodal understanding--work that could unlock new creative possibilities for millions of people.
Dive into the details!
https://t.co/1QNYkjWl4v
#NeurIPS2025 #AI
Adobe Research @ #NeurIPS2025!
Lots of exciting work will be presented this week. If you're interested in a deep discussion, potential collaborations, or internship opportunities, feel free to reach out β Iβm around as well π
papers: https://t.co/ugflkpj99O
π€ We rely on gaze to guide our actions, but can current MLLMs truly understand it and infer our intentions?
Introducing StreamGaze π, the first benchmark that evaluates gaze-guided temporal reasoning (past, present, and future) and proactive understanding in streaming video settings.
β‘οΈ Gaze-Guided Streaming Benchmark: 10 tasks spanning past, present, and proactive reasoning, from gaze-sequence matching to alerting when objects appear within the FOV area.
β‘οΈ Gaze-Guided Streaming Data Construction Pipeline: We align egocentric videos with raw gaze trajectories using fixation extraction, region-specific visual prompting, and scanpath construction to generate spatio-temporally grounded QA pairs. This process is human-verified.
β‘οΈ Comprehensive Evaluation of State-of-the-Art MLLMs: Across all gaze-conditioned streaming tasks, we highlight fundamental limits of current MLLMs. All MLLMs fall far below human performance. Models particularly struggle with temporal continuity, gaze grounding, and proactive prediction.
π World record performance: SambaNova is running Llama 3.1 405B at 114 t/s with full precision accuracy, in only one rack. Verified by @ArtificialAnlys! π¦
This speed unlocks so many use cases for enterprises and developers that we cannot wait to see them built on our platform.
Apply for early access today: https://t.co/CSlbJbTFVj
If you are at #NAACL2024 today, come by our poster 4οΈβ£9οΈβ£ to check out our explainable, **editable**, part-based image classifier!
Users can intervene the classification process or even modify the classifier by simply editing text descriptors.
Thanks for the shoutout + covering our #ACL2023nlp work on MeetingQA, @JayAlammar@cohere@forai_ml! It was great interacting with you π
PS: For those interested, details at π https://t.co/t706r2ZxLL
cc/ Trung, @david_s_yoon, Hanieh, @FranckDernoncou and @mohitban47
New #ACL2023nlp paper: π π²π²ππΆπ»π΄π€π!
We explore how good LMs are at answering questions in complex conversational meeting settings (with rhetorical + discussion seeking questions & multi-span + multi-speaker answers)
https://t.co/CjrG3CYTpT
@AdobeResearch @uncnlp
π§΅π
Adobe researchers are presenting new #ComputerVision work at this weekβs #CVPR2023. In addition to the publications, @Adobe authors have also contributed to the conference in many different ways. Check out the blog post to learn more! https://t.co/CarCfJr2QI
In the NYT today, Cade Metz implies that I left Google so that I could criticize Google. Actually, I left so that I could talk about the dangers of AI without considering how this impacts Google. Google has acted very responsibly.
Adobe researchers are presenting new work at #EMNLP2022, one of the top research conferences on #NaturalLanguageProcessing. Check out the blog post to learn more! https://t.co/ZxQebCGtnm
Check out the full list of @Adobe co-authored papers and other contributions at this year's @Siggraph, the premier academic conference in computer graphics and interactive techniques! https://t.co/prsqzffyPt