Lamb Labs (YC S26) is building custom AI inference chips that can do 20,000+ tok/s at 63x higher Intelligence per Watt (IPW) than traditional GPUs.
GPUs became the default for AI because they were available, not because they were the right hardware.
@amritessh re: hardcoding model weights, our bet is that models will start stabilizing plus for a lot of tasks, a lot of the models already do a good job, i.e., you donβt always need the frontier/SOTA modelβ¦but also automating the chip design pipeline!
Lamb Labs (YC S26) is building custom AI inference chips that can do 20,000+ tok/s at 63x higher Intelligence per Watt (IPW) than traditional GPUs.
GPUs became the default for AI because they were available, not because they were the right hardware.
A $100M GPU deal doesn't close on an exchange. It closes in a group chat, off a voice note.
Today we're launching Stoa: an RFQ marketplace for GPUs. $300M+ in RFQs in our first month.
Post what you need. Vetted dealers bid blind. Firm quotes within 48 hours. We handle KYB, contracts, shipping, and settlement.
Every RFQ, firm quote, award, and settlement runs through Stoa, so we see what hardware trades for. That history becomes prices buyers, sellers, and lenders can mark, finance, and hedge against.
@stoaexchange
Big news: Kimi-K3 by @Kimi_Moonshot is now #1 in the Frontend Code Arena with 1679 pts, surpassing Claude Fable 5.
This is a 17-place jump from Kimi-k2.6 (#18 -> #1).
In Frontend, Kimi-K3 ranked #1 in 6 of 7 domains: Brand & Marketing, Reference-Based Design, Data & Analytics, Consumer Product, Simulations, and Content Creation Tools, landing #2 only in Gaming behind Fable 5.
The full model weights will be released by July 27.
Congrats to the @Kimi_Moonshot team on this major milestone!
The team attended the #AUTONOMOUS2026 conference and had some great conversations around intelligence per watt, Physical AI, and bringing intelligence to the edge.
As Physical AI moves from demos to real world deployment, the hardware itβs running on matters.