🥳 Google is donating Agent Substrate to CNCF, and our submissions is accepted today! This effort has now the full potential to become the underlying compute runtime layer for all agentic systems. https://t.co/tE1bhdVryl
I've been hearing a lot about different programming workflows to make full use of LLMs, but I want in-depth accounts of how it works. This blog by @harper is exactly what I've been looking for.
https://t.co/mslbySfHla
After 6+ months in the making and burning over a year of GPU compute time, we're super excited to finally release the "Ultra-Scale Playbook"
Check it out here: https://t.co/dekxY4BQZO
A free, open-source, book to learn everything about 5D parallelism, ZeRO, fast CUDA kernels, how and why overlap compute & communication – all scaling bottlenecks and tools introduced with motivation, theory, interactive plots from our 4000+ scaling experiments and even NotebookLM podcasters to tag along with you.
- How was DeepSeek trained for $5M only?
- Why did Mistral trained an MoE?
- Why is PyTorch native Data Parallelism implementation so complex under the hood?
- What are all the parallelism techniques and why were they invented?
- Should I use ZeRO-3 or Pipeline Parallelism when scaling and what's the story behind both techniques?
- What is this Context Parallelism that Meta used to train Llama 3? Is it different from Sequence Parallelism?
- What is FP8? how does it compares to BF16?
In this book, our goal was to gather, in a single place, a coherent, easy to read yet detailed story of all the techniques that make today's LLM scaling possible.
The largest factor for democratizing AI will always be teaching everyone how to build AI and in particular how to create, train and fine-tune high performance models. In other word making accessible to everybody the techniques that power all recent large language models and efficient training is possibly one of the most essential of them.
What started as a simple blog-post ended up becoming an interactive writing piece containing 30k+ words. So we've decided to actually print it as a real 100-pages physical book as well: the physical ultrafast playbook –containing all the science of distributed and fast AI training.
We plan to send free copies as gifts to the first readers of the online version so feel free to add your email in the form linked in the blog post.
Ollama supports embedding models! Bring your existing documents or other data, and combine it with text prompts to build RAG (retrieval augmented generation) apps!
Learn more:
https://t.co/AbuiSytmDm
CVE-2023-45290 I've discovered has been fixed in this release. However, the misuse of the net/textproto.Reader in Go projects might still pose hidden risks. Check my article to learn more about this Golang issue and to avoid making similar errors: https://t.co/fPsJUTOoxq
What if we set GPT-4 free in Minecraft? ⛏️
I’m excited to announce Voyager, the first lifelong learning agent that plays Minecraft purely in-context. Voyager continuously improves itself by writing, refining, committing, and retrieving *code* from a skill library.
GPT-4 unlocks a new paradigm: “training” is code execution rather than gradient descent. “Trained model” is a codebase of skills that Voyager iteratively composes, rather than matrices of floats. We are pushing no-gradient architecture to its limit.
Voyager rapidly becomes a seasoned explorer. In Minecraft, it obtains 3.3× more unique items, travels 2.3× longer distances, and unlocks key tech tree milestones up to 15.3× faster than prior methods.
We open-source everything. Let generalist agents emerge in Minecraft! Welcome you all to try today: https://t.co/1d3YocozsI
Paper: https://t.co/JcWEasgtyI
Code: https://t.co/KsvVf7rcl0
Deep dive with me: 🧵
What should you do with your teams and systems after a period of hypergrowth? Listening to this @InfoQ talk from @pcalcado that offers lessons learned ... https://t.co/lg4vEMuU02
Extreme Programming provides an undo button, a game-changer for a single team - @KentBeck . But multiple teams need ways to manage inter-dependencies. I particularly like that he indicates the importance of Slack
https://t.co/EIkayoFdVF
@kelseyhightower I suspect the best form of verification on Fediverse is by having a known entity running the instance. Because my instance is run by Thoughtworks, my employer verifies my account. (I wrote more on this at https://t.co/Fi7lXuokSw)
🎉Great news!🎉
We're thrilled to announce that we have just open-sourced the Tgrade code🔥
🧪This permissive #OpenSource license enables #validators to run our #testnets and make any adjustments to their infrastructure.
Instructions and more info���
https://t.co/L5ExGPsVyC
Tailscale is fully remote. Making sure everyone gets a chance to speak and be heard in our internal meetings matters.
How do we do that? With a virtual red carpet of course! @commaok explains.
https://t.co/90J6SC5pYm
Between the 3 Sept and 10 Sept, secure env vars of *all* public @travisci repositories were injected into PR builds. Signing keys, access creds, API tokens.
Anyone could exfiltrate these and gain lateral movement into 1000s of orgs. #security 1/4
https://t.co/i23jFzAjjH
Looking to get started with #SRE:
👉New SRE site : https://t.co/LQ8L1l5JjD
👉5 resources to get you started : https://t.co/FHtjtFYZ3b
👉SRE Books : https://t.co/MGPr8FO4wf
���Profesional Cloud DevOps Engineer certification : https://t.co/fmNc537LRY