The amazing folks at @EdinburghNLP will be presenting a few papers at ACL 2025 (@aclmeeting); if you're in Vienna, touch base with them! Here are the papers in the main track 🧵
New Anthropic Research: “Inverse Scaling in Test-Time Compute”
We found cases where longer reasoning leads to lower accuracy.
Our findings suggest that naïve scaling of test-time compute may inadvertently reinforce problematic reasoning patterns.
🧵
🏆 Our @nvidia KV Cache Compression Leaderboard is now live!
Compare state-of-the-art compression methods side-by-side with KVPress. See which techniques are leading in efficiency and performance. 🥇
https://t.co/kP9fdEG5JZ
Hi! I will be attending #NAACL2025 and presenting our paper on self-training for tool-use today, an extended work of my MSc dissertation at @EdinburghNLP, supervised by @PMinervini.
Time: 14:00-15:30
Location: Hall 3
Let’s chat and connect!😊
NAACL 2025 Oral Presentation💥
Our work about using Sparse AutoEncoder to resolve knowledge conflict will present on 30 Apr 11:30–11:45 AM • Ballroom C
Thank Hongru for presenting our work!!!
Grateful to be part of this! I'll be presenting our paper "Self-Training Large Language Models for Tool-Use Without Demonstrations" at #NAACL2025! Also, I am currently seeking PhD opportunities. Please feel free to reach out if you're recruiting or know of any openings! :)
My amazing collaborators will present several works at ICLR and NAACL later this month -- please catch up with them if you're attending! I tried to summarise our recent work in a blog post: https://t.co/wF8dqldSqv
We find a single biased direction encodes a KV Cache selection mechanism in Self-Attention -- Key vector with a strong component in this direction results in this Key-Value pair being ignored by Query🚀🚀🚀