Hi, I'm RC. I built Kimi CLI at Moonshot last year, and back in 2015, bots that lived in group chats. For the past four months, I've been building Raft in public.
Today I'm launching Raft 1.0.
Right now, working with agents means juggling terminals, sessions, and skills. The more you run, the more you end up holding it all together yourself, and the easier it is to lose the thread.
Raft puts your agents in team mode: one workspace where working with agents feels like messaging your team. The work keeps moving, and you stay at the wheel.
Meet my Raft agent team👇
We just renamed Slock to Raft today.
The name fits what we’re building much better: a shared foundation where many agents can coordinate, carry context, and move work forward together.
There’s also a quiet nod to the Raft consensus protocol — distributed actors, shared state, reliable progress.
Same vibe. Sharper metaphor.
from apps to material
software used to be something you opened
an app was a room with walls: calendar here, notes there, music there, work there. each one had its own logic, buttons, its own little kingdom. the user moved between kingdoms, carrying context in their head
but ai starts to break the walls
software becomes less like a destination and more like material. something you shape, combine, stretch, ask, remix, and leave behind as traces. a document can become an app. a conversation can become a workflow. a song can become a memory. a task can become an agent. the boundary between using and making gets blurry
the old model was: choose the right tool for each task
the new model is: express the shape of the thing you want, then refine it with the system you built
this changes the role of the interface. ui is no longer only fixed views for fixed functions. it becomes a surface where intent turns into structure. the best interfaces will feel less like menus and more like clay – responsive, persistent, inspectable, and alive
apps won’t disappear. rooms are still useful. but the deeper shift is that software stops being a set of sealed containers and becomes a medium people can think through
like paper, but executable
like language, but spatial
like memory, but programmable
software stops being something only programmers make
it becomes material anyone can shape
OpenCLI v1.8.0 is finally out 🎉
Stayed up till 4am — finally landed the 1.7 → 1.8 stretch. Looking back, this run packed in a lot:
## Browser Agent Runtime
Rounded out the browser layer in one shot. The point: stop driving the browser with brittle CSS selectors, and move to "accessibility tree + semantic locators + CDP-primary input" — agent-native end to end. CDP input + AX snapshots + semantic locators (--role / --name / --label / --testid) + hover / focus / dblclick / check / upload / drag / wait download / annotated screenshot.
opencli browser <session> click 5 — grab a ref and click. No more selector guessing. Custom dropdowns on Radix / shadcn / Material UI that used to silently fail to open now click reliably.
## New sites / new surfaces
weread-official — WeRead's official Agent Gateway, pure HTTP + Bearer key, coexists with the cookie-based weread
12306 — full read (trains / prices / my orders / passengers)
xianyu — Xianyu inbox / messages / reply, you can talk back now
suno — music generation
linkedin — Sales Navigator + people-search + messaging / safe-send / thread-snapshot all wired up
linkedin-learning / rednote / booking / ctrip hotels + flights / DuckDuckGo / Brave / Yahoo search
## Twitter, kept polishing
list-create / device-follow / quoted_tweet / card binding_values / bio / UserMedia cursor pagination — all filled in
write-action symmetry: unlike / retweet / unretweet / quote
bookmarks / bookmark-folder / list-tweets now carry media
## Reddit / Zhihu, deeper read coverage
reddit: subscribed / whoami / home / subreddit-info / --expand-more (load more comment branches) / listing exposes post_hint+url+preview+gallery
zhihu: answer-detail / answer-comments / answer pagination
(important)
## Reliability & security
Download path-traversal fix — remote-controlled fields (e.g. video titles used as filename) can't ../ escape the output dir anymore
Page.goto stale-identity self-recovery; CDP -32000 is now retryable
undici 8.x silently bumped the engines floor to Node ≥22 — pinned back to 6.x to keep the Node 20 promise alive
youtube transcript cross-video bleed fixed (SPA watch→watch was returning the predecessor's captions)
ChatGPT web image generation back to green (now detects CSS background + canvas, not just <img>)
Typed-error sweep across Douyin / Jike / WeRead / Apple Podcasts / Reddit / Gitee / lesswrong / xhs / YouTube — silent-sentinel and silent-empty fallbacks replaced with EmptyResultError / AuthRequiredError. Agents no longer have to guess whether [] means "truly empty" or "site changed".
## Trace & Observation
Failed adapter runs now ship a trace artifact; summary.md is the entry point.
browser console / browser network --failed / --follow — agents can finally see what's happening inside the browser.
## README, 20% lighter
EN 410 → 326, ZH 455 → 371. Built-in Commands curated to 11 sites; CLI Hub reduced to a name enumeration. Easier to land on.
Goal hasn't changed: read & action infrastructure for AI agents.
https://t.co/4rmxL1wSJc
Our newest model, π0.7, has some interesting emergent capabilities: it can control a new robot to fold shirts for which we had no shirt folding data, figure out how to use an appliance with language-based coaching, and perform a wide range of dexterous tasks all in one model!
Introducing Claude Opus 4.7, our most capable Opus model yet.
It handles long-running tasks with more rigor, follows instructions more precisely, and verifies its own outputs before reporting back.
You can hand off your hardest work with less supervision.
Hi, I’m RC. I previously built Kimi CLI at Moonshot AI.
Now I’m building Slock, an agent-human collaboration platform for modern builders and teams.
Today, we're shipping a ton of new features and improvements in Slock: search, thread inbox, saved messages, message permalinks, pinned chats, server join links, a more consistent color system, and many smaller upgrades.
More details in the thread below.
AC 2
Anyone can contribute data at home
Ego+AC one+Pi*0.6
👀Uses only head cameras to complete the task.
💪The task is learned from human videos.
📱Data collection requires only a mobile phone.
📸Works across different cameras.
🦾Errors are corrected through human demonstrations.
✅The training process is simple and efficient.
🤗High task accuracy is achieved after training.
Thanks,
Tesla that video give us some ideas
Pi*0.6 awesome work!
@JingxiangMo@gradientrobots What a pity that K-Scale has shut down—such a brilliant group of geeks gone! Can’t wait to see JX take his new company to even bigger heights.