Check out our new work on combining General Preference Modeling (GPMs) with GRPO style methods. Our proposed GRPL algo prevents reward hacking and improves reasoning quality in many open-ended tasks compared to the baselines.
My contributions were only advisory on this project.
Checkout our new work on exploring different inductive biases of causal vs bi-directional language modeling objectives vis-ร -vis latent generalization (reversal curse). This work was led by @juliancodaforno as a student researcher in @GoogleDeepMind
๐ Excited to share our new paper from my SR work @GoogleDeepMind : "The Illusion of Latent Generalization: Bi-directionality and the Reversal Curse" . It was accepted at the ICLR 2026 Workshop on Representational Alignment (Re-Align)!
A few colleagues and friends are organizing a workshop at ICML in Seoul on the continuous adaptation of frontier models. Consider submitting your works to the workshop and/ or attend it if you are going to ICML.
Link: https://t.co/8yAblYtyLA
[8/8] There are a lot of other interesting details and experiments in the paper. Check it out: https://t.co/bYt6K2rSI3) and provide us with any feedback.
Joint work with @Sridhaar96@AndrewLampinen
[1/8] Our new paper on using test-time compute to improve latent learning: https://t.co/bYt6K2rSI3
Thinking models have been shown to improve performance on maths, reasoning and coding tasks. In this work, however, we look into how thinking models improve latent learning.
[7/8] Yet, pure reversals lack intermediate reasoning chains, making them difficult for thinking models; here, ICL remains the best baseline. This suggests scaling compute alone cannot solve pure reversals, implying a need for architectural changes.
@haider1 When pondering about the effects of technological revolutions on the labor market, listen to economists who have studied the effects of past technological revolutions on the labor market.
Never listen to computer scientists.
Iโm excited to share our new @Nature paper ๐, which provides strong evidence that the walkability of our built environment matters a great deal to our physical activity and health.
Details in thread.๐งต
https://t.co/omO3YcHrvG
[1/4] I am hiring a student researcher to work on topics related to continual learning, knowledge acquisition, and science of learning in LLMs this summer. This will be an in-person position and you will be based in Mountain View. https://t.co/jlgdZelgQG
[3/4] If you are interested, please send me your resume, brief snippet of your research, and what you would want to explore during the internship at [email protected]. Add [SRP 2025 Candidate] in your email subject.