So basically there's this guy Andrew. He posted a rumor about Anthropic solving navier stokes. Now before we go ahead let's go back in time.
There's this guy levent, he's a mathematical hunk of a dude, makes these cute posts about solving the jacobian conjecture like he just cooked ramen.
Anyway this dude approached Tristan Buckmaster, a math professor from NYU, to collab with him a year ago to solve navier stokes.
They made huge progress using Anthropic and OpenAI models but wanted to release the paper in a normal manner that math nerds could read instead of AI slop.
In the meantime OpenAI was rife with rumors about this stuff. So the prof reached out to a guy he knew at OpenAI to compare notes and address rumors.
According to Tristan he was then rushed into a meeting with his contact and Sebastien Bubeck (smart guy, some say he's the smartest guy at OpenAI, not me but a lot do)
Turns out they both were solving the same problem. Prof then asked how they were doing it and found that they were taking the same path as him which he found hella shady (could be shady, could be big computer who knows)
Prof had used OpenAI products to do research so he asked OpenAI if they looked up his user data and they said no (believable would lead to lot of lawsuits if not) but then prof hit them with the counter about it being trained on that data and got no response.
He did not get an answer but got two proposals (very godfathereqsue)
> he posts his result, openai post theirs but say that he deserves the clay prize (million dollars as this is one of em millennium problems) and that they were the humans closest to it
> he posts his result and the openai result but removes levent coz he's an anthropic boy (OpenAI and Anthropic are like tupac and biggie or drake and kendrick)
Prof declined both and threatened to go public so Sebastien asked why he would ruin his career like that?
Prof got testy and asked why it would ruin his career and Sebastien called him irrational (an emotional bitch when translated from SV to common parlance) behind his back and asked levent to meet secretly who said no (why would he not say no man come on?)
Then prof went public and threw enormous shade on OpenAI though he claimed he didn't have all the facts (but he implied they stole stuff heavily)
Sebastien then posted on twitter to defend his honor saying the speculations have no merit.
Meanwhile another Anthropic guy Sholto Douglas (designated anthropic podcast guy and Australian) told us that levent is a very sweet guy and he was sad he couldn't present his year long collab work they way he wanted (a single tweet and an AI proof probably)
OpenAI top dog Noam Brown thought it would be funny to parody this by copying Sholto but quoting Sebastien and saying he was sad he couldn't present his "week" long collab work they way he wanted. Don't know if he was trying to help or bury the guy, very shakespearean.
Now both openAI and anthropic guys are tweeting about being best buddies and solving alignment. Though openAI are tweeting officially tomorrow.
Most people have no idea that their Macbooks already has powerful AI models running locally.
I'm building something that puts them to use.
Stay tuned. ๐
@aclmeeting KG-MuLQA turns financial docs into KG-based QA pairs to test multi-hop retrieval, set ops, answer plurality, implicit relations etc.
paper: https://t.co/xYunNKlVF5
curious how others separate โread itโ from โreasoned over itโ?
super excited that KG-MuLQA, a paper i co-authored, is accepted as a main conference paper at @aclmeeting 2026 ๐ฅณ
i wonโt be at ACL this year, but would love to talk evals / benchmarks / long-context llms. dms open :)
#ACL2026#LLM
iโll be presenting 2 solo workshop papers:
PEBS: per-rater reward model calibration @ Pluralistic Alignment
https://t.co/TxG1SjLr0P
RAC: delayed feedback in RLHF at RLxF
https://t.co/6JWKj9jET8
RAC also got a nice summary by @gistdotscience:
https://t.co/1SxDrEfTJa
iโll be at #ICML2026 in seoul july 6-11 ๐ฐ๐ท
first time presenting my own work around an A* conf, so this feels very special.
would love to meet rl/rlhf folks, reward modeling people, alignment folks or anyone down for coffee, korean bbq, or late-night chimaek.
dms open :)
this is mine too :)
PEBS came from a very simple reward modeling problem: annotators donโt all use rating scales the same way.
if we pool everyone into one reward model, we can end up fitting an โaverage raterโ who doesnโt really exist.
#ICML2026#RLHF ๐ฐ๐ท
this is mine :)
RAC came from a rl/rlhf issue i kept running into while making rl envs for agents: good feedback is often slow.
it tries to use delayed feedback retroactively instead of waiting for it or dropping it.
see you at #ICML2026 ๐ฐ๐ท
Retroactive Advantage Correction: Closed-Form V-Trace Bias Correction for Delay-Aware RLHF
Arnav Raj
https://t.co/dWG31D8dAS [๐๐.๐ป๐ถ ๐๐.๐ฐ๐ธ]
๐ฌAccepted at the ICML 2026 Workshop on Reinforcement Learning from World Feedback (RLxF)
whoโs coming to ICMLโ26 ๐ฐ๐ท?
making a very unserious-but-useful whatsApp group for the ICML side quests:
Seoul food runs, travel plans, research chats, parties, and people to hang out with between sessions.
join here: https://t.co/G62dpdG6tZ
share with friends coming too :)
whoโs coming to ICMLโ26 ๐ฐ๐ท?
making a very unserious-but-useful whatsApp group for the ICML side quests:
Seoul food runs, travel plans, research chats, parties, and people to hang out with between sessions.
join here: https://t.co/G62dpdG6tZ
share with friends coming too :)
whoโs coming to ICMLโ26 ๐ฐ๐ท?
making a very unserious-but-useful whatsApp group for the ICML side quests:
Seoul food runs, travel plans, research chats, parties, and people to hang out with between sessions.
join here: https://t.co/G62dpdG6tZ
share with friends coming too :)
@heiga_zen@irvinxyz Do you accept international students? I am doing masters at IIT Delhi, I have 2 solo paper in rlhf at ICML workshops, 1 ACL main paper and iclr workshop paper.
More about me: https://t.co/lltGE4umb1
whoโs coming to ICMLโ26 ๐ฐ๐ท?
making a very unserious-but-useful whatsApp group for the ICML side quests:
Seoul food runs, travel plans, research chats, parties, and people to hang out with between sessions.
join here: https://t.co/G62dpdG6tZ
share with friends coming too :)
whoโs coming to ICMLโ26 ๐ฐ๐ท?
making a very unserious-but-useful whatsApp group for the ICML side quests:
Seoul food runs, travel plans, research chats, parties, and people to hang out with between sessions.
join here: https://t.co/G62dpdG6tZ
share with friends coming too :)