Most of my past year went into leading the inference efforts for this model. I’m a strong believer in frontier intelligence that’s efficient and usable across devices. Incredibly grateful to have worked with this team. Congrats @deepgrove_ai!
We introduce Maple-Preview, an open-source 20B-A1B ternary-weight reasoning LLM, SOTA in its weight class.
It solves IMO-level problems and runs at 200+ tokens/s on a Mac Mini M4, 5–16× faster than efficient models like Gemma 4, Qwen3.5, and gpt-oss.
https://t.co/l2y5eZTVS4
We’re hiring!
The shift away from today’s monolithic AI generalist to a world where everyone truly owns their own model requires models to be efficient enough to run anywhere and adapt in real time.
We’re a small team obsessed with pushing this frontier. We’re scaling pre-training, post-training, and infra for efficient yet strong low-bit models.
If this sounds like you, we'd love to chat: https://t.co/MI3MZYtS6b
Conquering the next battlefield: Space.
We bet our own money to rapidly design, build, test, and fly a fully integrated satellite. Last month, ANDURIL-216 reached orbit. Anduril, alongside a few core partners, went from clean sheet to on orbit in less than two years.
All systems are nominal. We've made first contact and fully commissioned the bus, power distribution devices, and edge compute payload. The first calibrated images are in from all three payloads onboard the spacecraft: one long-wave infrared, two electro-optical.
The firsts go beyond the vehicle. Our space team used Lattice to commission and task ANDURIL-216, the first time Lattice has commanded an on-orbit spacecraft.
More to come.
Most of my past year went into leading the inference efforts for this model. I’m a strong believer in frontier intelligence that’s efficient and usable across devices. Incredibly grateful to have worked with this team. Congrats @deepgrove_ai!
We introduce Maple-Preview, an open-source 20B-A1B ternary-weight reasoning LLM, SOTA in its weight class.
It solves IMO-level problems and runs at 200+ tokens/s on a Mac Mini M4, 5–16× faster than efficient models like Gemma 4, Qwen3.5, and gpt-oss.
https://t.co/l2y5eZTVS4
We introduce Maple-Preview, an open-source 20B-A1B ternary-weight reasoning LLM, SOTA in its weight class.
It solves IMO-level problems and runs at 200+ tokens/s on a Mac Mini M4, 5–16× faster than efficient models like Gemma 4, Qwen3.5, and gpt-oss.
https://t.co/l2y5eZTVS4
Introducing Rook by @Hop_Aero (YC S26): an autonomous hypersonic cargo rocket that delivers 550 lbs up to 450 miles in ~15 minutes.
It launches from a standard 40-foot shipping container and lands on unprepared surfaces—no runways or fixed infrastructure required.
not sure if it's just me, but I feel people should report inference perf as
1. Prefill: input tok/s/GPU at ctx len
2. Decode: output tok/s/GPU at ctx len
3. KV cache size
Not total tok/s/GPU, which varies a lot. Feels like marketing numbers, really hard to compare.
It was remarkable to watch @anduriltech's first Ohio-made, unmanned combat aircraft roll off the production line at Arsenal-1 in Pickaway County this morning. Congratulations to everyone involved in this significant milestone for national defense!
Introducing Waddle Labs: Claude Code for robots.
Connect our API to your robot and enter a prompt, then our agents write code to achieve the task in 20 minutes.
@yiding_song@theWaddleLabs