Top Tweets for #colm2026
๐ Outstanding Paper Award! @WengZhaoti39773 @xwang_lk
Proud to share that โGroup-Evolving Agentsโ received the ๐๐ฎ๐ญ๐ฌ๐ญ๐๐ง๐๐ข๐ง๐ ๐๐๐ฉ๐๐ซ ๐๐ฐ๐๐ซ๐ at the Lifelong Agents Workshop @ #COLM2026 ๐
Worth all the hard work, and well deserved! Canโt wait to see whatโs next!
๐ Honored that Group-Evolving Agents won the Outstanding Paper Award at the Lifelong Agents Workshop @ #COLM2026!
Agents that evolve as a group, sharing experience, keep improving open-endedly ๐งฌ๐ค
Grateful to my wonderful coauthors and @xwang_lk !
Had a great time at #COLM2026
๐ https://t.co/f1JVRsrwqD
๐ป https://t.co/ckntpuTgaR

๐ Honored that Group-Evolving Agents won the Outstanding Paper Award at the Lifelong Agents Workshop @ #COLM2026!
Agents that evolve as a group, sharing experience, keep improving open-endedly ๐งฌ๐ค
Grateful to my wonderful coauthors and @xwang_lk !
Had a great time at #COLM2026
๐ https://t.co/f1JVRsrwqD
๐ป https://t.co/ckntpuTgaR

Enjoyed talking with Yuandong @tydsh!
@Recursive_SI @COLM_conf
Check our latest work on RSI and CUA as well!
https://t.co/Vi43CBmZy3
https://t.co/YWo4a6tLjR
https://t.co/tbRw2hD7nC
#RSI #Recursive #COLM2026 #ClawBench #WebsiteBench #RewardHarness #SelfImprovement

That's a wrap on #COLM2026! Thank you to everyone who stopped by our booth, joined our evening gathering, & talked open research with us.
Catch up on the papers, interviews, & highlights from the week. ๐
https://t.co/vou0wAYbfB
Zora is presenting this (and her other awesome work) as an invited talk at LSEI now - Union Square 15&16! #COLM2026

Agents trained on population-scale data can do a lot of things.
But ask a professional to stake their reputation on an AI-generated artifact? โPretty goodโ isnโt good enough.
We introduce TAHI: a Test-time Adaptive agent framework through Human-agent Interaction. Featuring:
๐ Test-time adaptation via context (memory, skills) and weight training
โก Efficient adaptation to individual expertise within tens of tasks
โ๏ธ Creating comprehensive rubrics for โnon-verifiableโ tasks
๐ Analysis of shared community guidelines vs. personalized tacit expertise
Heading back from lunch saw 4 Waymos in a row ๐
#COLM2026

Golden Gate again
#colm2026

At the poster, Franciscan C and D, until 1:55. The rule: a memory enters the prompt only if everyone listening now was there when it was learned. Add one person to the room and it knows less. That's the feature. #COLM2026
What I read at #COLM2026: the 5 papers I liked most, ranked by community likes:
1. IdeaScientist: agents trained with RL to generate grounded research ideas. @Jiarui_Liu_
https://t.co/CM0N8qBqgI
2. LLM-as-a-Verifier: a weaker model checking a stronger oneโs work. @Azaliamirh
https://t.co/1I6P5e63kI
3. CORAL: autonomous multi-agent evolution for open-ended discovery. @ao_qu18465
https://t.co/yitgWGuYVA
4. Actor-Curator poster. @lightetal
https://t.co/yuYkafFV1Y
5. GitSwarm: inference that builds on itself across long tasks. @veds_12
https://t.co/hhIJh0tA2t
My overall impression is that verification is now the bottleneck for both agents and reasoning. Which COLM paper did I miss?
Excited to share IdeaScientist, our research from my internship at Meta! ๐
Can AI agents learn to generate novel, grounded scientific ideas?
We introduce IdeaScientist, a framework that trains specialized agents to identify research gaps, discover useful connections across scientific domains, and turn them into concrete research proposals.
๐ Key findings:
โข +14% overall performance over open autoresearch baselines
โข 25% improvement in novelty, showing that explicit training can improve scientific ideation
โข 75โ92% human preference win rates against six open baselines in blind evaluations
โข Cross-domain retrieval substantially improves the transfer of ideas between research fields
We also introduce Svalbard Idea Vault, a collection of 2.77M decomposed research ideas for training and evaluating scientific ideation.
๐ Paper: https://t.co/OCFIyCKe4k

@i_beltagy Great bringing together friends and collaborators from industry nonprofits and academia at the Ai2 social yesterday different goals all aligned on open source AI #colm2026
Come checkout our poster SUPERNOVA: Eliciting general reasoning in LLMs with Reinforcement Learning on Natural Instructions!!!
TL;DR: instruction datasets are an underused RLVR source, but utilizing them requires principled empirical curation.
๐ Imperial Ballroom B #COLM2026

Reminder: our final schedule for COLM 2026 Workshop on Efficient Reasoning

Excited to be at #COLM2026 today! Happy to chat about scalable LLM data processing, deep research and our latest paper on DeepScholar-bench.
If youโre building and benchmarking DR, Negar and I would love to chat :)
Had a great time presenting DeepScholar-Bench at #COLM2026 today! Lots of interesting conversations and interactions! ๐คฉ
If youโre building a deep research system, give our benchmark a try!
@COLM_conf has been such an engaging and fun conference so far! ๐

Great to connect with folks over the past few days in SF!
Paper: https://t.co/CnVGy89v6i
#COLM2026 #MathReasoning #LLM #AIEvaluation (3/3)
A model reads a proof with a gap. Instead of flagging it, it invents an argument to fill the gap and calls the proof correct.
That's one failure mode we found using small open models as proof judges. Our last #COLM2026 paper: (1/3)
Do LLMs need to reprocess an entire long document for every question?
HยฒMT builds reusable latent memories around document structure and prunes irrelevant branches early.
๐ Presenting today at the #COLM2026 Context Beyond the Window workshop. Stop by!
https://t.co/abnPjH9m6o
Interested in new architectures & unlearning that actually works? Come checkout Gaurav's oral at HAIPS workshop at #COLM2026. Natively unlearnable language models are a pretty cool idea!
@_christinabaek @datologyai @gaurav_ghosal @AdtRaghunathan NULLs (Natively Unlearnable LLMs) thread
https://t.co/FQea9M9e60

Check out our @GenAI4World workshop at COLM today!
We have an exciting speaker lineup and set of paper presentations, and I'll be moderating a cool panel at the end of the day!
https://t.co/ymFgbqaLm3
Can you fingerprint agents from a few turns of non-adversarial conversation without access to system prompts or model internals?
Can we trace scammers and undisclosed API changes?
๐10:10 - 10:35 and 3:25-4:15 at HAIPS at #COLM2026 Square 19 & 20

Presenting today at the LLM/VLM Deployment Opportunities and Risks in Healthcare workshop at #COLM2026! ๐ง ๐ฉบย
Stop by our poster to chat about:
๐ MindEval: Benchmarking Language Models on Multi-turn Mental Health Supportย (https://t.co/QnX0ndaGe5)
๐ Union Square 23&24 (Fourth Floor)
๐๏ธ Today | Workshop Session
Last Seen Hashtags on Sotwe
profesora
Seen from Mexico
ูุญุณ_ุงููุณ
Seen from Germany
ifลa sarhoล
Seen from Turkey
minichat(**)filter:native_video
Seen from United States
lucayu
Seen from United States
RevolverRevolver
Seen from United States
teenageee
Seen from Germany
HugeMed
Seen from United States
ุฏููุช_ุฃุฎุชู
Seen from United States
Neighbor
Seen from India
Most Popular Users

Elon Musk 
@elonmusk
241.5M followers

Barack Obama 
@barackobama
118.9M followers

Cristiano Ronaldo 
@cristiano
114.6M followers

Donald J. Trump 
@realdonaldtrump
111.8M followers

Narendra Modi 
@narendramodi
107.1M followers

Rihanna 
@rihanna
98.7M followers

NASA 
@nasa
92.3M followers

Justin Bieber 
@justinbieber
91.8M followers

KATY PERRY 
@katyperry
90M followers

Taylor Swift 
@taylorswift13
84M followers

Lady Gaga 
@ladygaga
75.5M followers

Virat Kohli 
@imvkohli
73.5M followers

Kim Kardashian 
@kimkardashian
70.9M followers

YouTube 
@youtube
68.8M followers

Neymar Jr 
@neymarjr
66.5M followers

Bill Gates 
@billgates
65.2M followers

Selena Gomez 
@selenagomez
63.1M followers

The Ellen Show
@theellenshow
62.2M followers

CNN 
@cnn
61.8M followers

X 
@x
60.7M followers































