Kimi K3 has received far more love than we expected, and our GPUs are feeling it.
Over the past 48 hours, demand has pushed close to the limits of our current capacity. To protect the experience of existing subscribers, we're temporarily pausing new subscriptions and prioritizing compute for current members. Existing subscribed users are not affected.
We're adding capacity as fast as we can and will reopen new subscription spots in batches.
Going forward, we'll also split membership into two more focused plans: Kimi Membership for Kimi Web, App, and Work; and Kimi Code Membership for coding workflows. This will help us match compute more precisely and keep the experience stable.
Thank you for your patience and understanding!
We got a new HTTP method before GTA 6. 💀
After decades of the same core HTTP verbs, we're finally getting a new one: QUERY.
It was published as RFC 10008 on June 15th 2026.
I am not a big fan of doing this client-side, for a whole slew of reasons, but this is impressive engineering by the @Zalando team. I like the "lessons learned" they conclude with. https://t.co/5kZkLsT2ro
Pewd did it again. now he open-sourced a self-hosted AI workspace. bro is building a CV harder than a CS undergrad looking for a job:
> built a 10-GPU home rig
> quantized giant LLM to run local
> built ChatOS, local AI UI
> added RAG/local memory
> built “council” of AI models
> built “swarm”, small models in parallel for data collection
> fine-tuned a Qwen 32B-based coding model
> donated compute from his GPU rig for protein folding research