Honored to be part of the OpenRSI @openrailwaymap effort with our team @ZZZZZPotentials !
Excited to help build open, rigorous evaluations of whether AI can improve not just a model, but the research process that produces the next one.
As AI systems enter a recursive self-improvement loop, the central question is whether it can systematically move beyond human-designed methods to genuinely extend the scientific and intelligence frontier.
What’s missing is a neutral, open standard for evaluating these capabilities in real, production-scale intelligence development.
Today, we’re releasing OpenRSI-Index v0.1: evaluating whether AI can recursively improve itself and push the boundaries of intelligence and science, on production-scale clusters with 1k GPUs.
We turn fully open-source projects into autoresearch environments, with agent trajectories lasting 60+ hours. Building v0.1 took 100K+ H100-hours.
We’re building an ecosystem with and for the research community: let RSI benefit everyone, and let everyone shape RSI together.
We invite task contributors and compute partners to build this open benchmark with us - all contributors will be included as paper authors.
Shape RSI with us:
🌐 Website: https://t.co/unaB2yfQy4
🛠️ GitHub: https://t.co/vaRyltgbra
🤝 Contribute: https://t.co/PffRAhoIht
1/ Blog
AI progress is limited not only by compute and data, but by the bandwidth of human insight.
The core ambition of RSI is to alleviate this bottleneck: by allowing agents to scale hypothesis generation and implementation, we can sweep a vastly larger method space than humans can explore alone.
Evaluating whether these loops can systematically surpass current methods is exactly what OpenRSI-Index is built for.
Research is never zero-sum: as agents handle the heavy lifting in the research loop, researchers are super-leveraged and get more room to chase crazier, paradigm-shifting ideas.
📖 Full write-up: https://t.co/VT8jmZITr5
Hi, I’m RC. I previously built Kimi CLI at Moonshot AI.
Now I’m building Slock, an agent-human collaboration platform for modern builders and teams.
Today, we're shipping a ton of new features and improvements in Slock: search, thread inbox, saved messages, message permalinks, pinned chats, server join links, a more consistent color system, and many smaller upgrades.
More details in the thread below.
We’re excited to share some updates on Genesis since its release:
1. We made a detailed report on benchmarking Genesis's speed and its comparison with other simulators (https://t.co/Wkr7gJtAGh)
2. We’ve launched a Discord channel and a WeChat group to foster communications between users and contributors
3. We released Genesis 0.2.1 today with new features including faster cached kernel loading, docker file support, smoke simulation, RL training example for drones, multilingual documentation support, together with various new APIs. A heartfelt thank you to the open-source community for contributing to this collaborative effort!
1/2 We kept hearing about @LanceDB from companies like MidJourney that were using its open-source Lance table format for AI data. They saw orders of magnitude improvement in latency, storage costs, and more. A quick experiment confirms why: https://t.co/Ojndl1kEvk.