We've released Run:ai Model Streamer as open source to help the community load inference models faster than ever before 🎉
This is part of our efforts at @runailabs to make inference workloads more dynamic and cost efficient.
GitHub repo:
https://t.co/BrOK2Chi6m
How Run:ai Streamer helps?
- 6x faster model loading times
- Works seamlessly across all storage types (local, S3, NFS)
- Native integration with popular inference engines like @vllm_project
- Zero model format conversion needed to allow seamless integration with @huggingface
I had the pleasure of being a guest on a couple of great podcasts recently 🙌
Shared thoughts about what we’re doing at @runailabs and about the advancements happening in the AI and GPUs world
Check out the episodes linked below
Discover the challenges and best practices of deploying LLMs into production environments.
@ronen_dar, Co-founder @runailabs, explores optimizing AI model training to reduce costs, facilitating the adoption across diverse business scales, and more.
🔗👇 https://t.co/Jx4TaSF6lC
Standing room only for @ronen_dar
of Run:ai at #kubecon#cloudnativecon co-located talk on training #LLMs on #Kubernetes. #AI needs dynamic compute but it’s hard to have GPU access when needed when GPU is shared manually.
@sharongoldman writes about how @runailabs and other Israeli AI startups are carrying on after the horrors of the last few days.
This is the biggest disaster Israel has ever had but as mentioned in the article “we’ll prevail and become stronger” 🇮🇱
https://t.co/SQLzF6K6JF
"Given that GPU availability is poised to remain a challenge for the foreseeable future, product leaders must think strategically about GPU allocation."
✅
Teams in Meta fight internally on access to GPUs
We’ve heard this story countless times from Run:ai’s prospects 🤦♂️
Apparently, GPU allocation is a huge problem even in “GPU-rich” companies
https://t.co/3ZZmMArRNb
Scheduling distributed computing workloads on shared CPU Clusters is not trivial
More about Run:ai’s Kubernetes CPU scheduling over here:
https://t.co/DVUFT3vWjT
Exclusive:
OpenAI is blowing by earlier revenue projections, already passed $1b ARR.
Turns out plenty of big companies like LLMs :)
https://t.co/4ei2UMIqJx w / @aaronpholmes@OpenAI $msft