AlphaFold 3 model code and weights are now both available for academic use on GitHub.
The golden age of computational science is about to begin!
https://t.co/sKiQbSS1jM
Starting today, open source is leading the way. Introducing Llama 3.1: Our most capable models yet.
Today we’re releasing a collection of new Llama 3.1 models including our long awaited 405B. These models deliver improved reasoning capabilities, a larger 128K token context window and improved support for 8 languages among other improvements. Llama 3.1 405B rivals leading closed source models on state-of-the-art capabilities across a range of tasks in general knowledge, steerability, math, tool use and multilingual translation.
The models are available to download now directly from Meta or @huggingface. With today’s release the ecosystem is also ready to go with 25+ partners rolling out our latest models — including @awscloud, @nvidia, @databricks, @groqinc, @dell, @azure and @googlecloud ready on day one.
More details in the full announcement ➡️ https://t.co/hhJoLm5eLV
Download Llama 3.1 models ➡️ https://t.co/rRjvmxqCTC
With these releases we’re setting the stage for unprecedented new opportunities and we can’t wait to see the innovation our newest models will unlock across all levels of the AI community.
FlashAttention is widely used to accelerate Transformers, already making attention 4-8x faster, but has yet to take advantage of modern GPUs. We’re releasing FlashAttention-3: 1.5-2x faster on FP16, up to 740 TFLOPS on H100 (75% util), and FP8 gets close to 1.2 PFLOPS!
1/
A theory of why Claude 3.5 Sonnet is insane at coding: mechanistic interpretability.
Anthropic showed that there are clever ways to understand what the weights of LLMs do and "steer" them to behave differently.
Doing this on Sonnet may be why it crushes it at code:
🧵
1/12
What if we could universally recombine, insert, delete, or invert any two pieces of DNA?
In back-to-back @Nature papers, we report the discovery of bridge RNAs and 3 atomic structures of the first natural RNA-guided recombinase - a new mechanism for programmable genome design
We have trained ESM3 and we're excited to introduce EvolutionaryScale.
ESM3 is a generative language model for programming biology. In experiments, we found ESM3 can simulate 500M years of evolution to generate new fluorescent proteins.
Read more: https://t.co/iAC3lkj0iV
Today, we’re thrilled to announce the open weights for Stable Diffusion 3 Medium, the latest and most advanced text-to-image AI model in our Stable Diffusion 3 series!
This new release represents a major milestone in the evolution of generative AI and continues our commitment to democratising this powerful technology.
🎉 Learn more and get started here: https://t.co/iZCpcV248M
I am so excited that xLSTM is out. LSTM is close to my heart - for more than 30 years now. With xLSTM we close the gap to existing state-of-the-art LLMs. With NXAI we have started to build our own European LLMs. I am very proud of my team. https://t.co/IH7giCe3gd
Schedule-Free Learning
https://t.co/D6OUDuJ05C
We have now open sourced the algorithm behind my series of mysterious plots. Each plot was either Schedule-free SGD or Adam, no other tricks!
Introducing Sora, our text-to-video model.
Sora can create videos of up to 60 seconds featuring highly detailed scenes, complex camera motion, and multiple characters with vibrant emotions.
https://t.co/YYpOAcrXQ3
Prompt: “Beautiful, snowy Tokyo city is bustling. The camera moves through the bustling city street, following several people enjoying the beautiful snowy weather and shopping at nearby stalls. Gorgeous sakura petals are flying through the wind along with snowflakes.”
Can finally talk some GPU numbers publicly 🙃
By the end of the year, Meta will have 600k H100-equivalent GPUs.
Feel free to guess what's already deployed and being used 😉!
the 18-year-old hacker who leaked GTA 6 clips has been sentenced to life in a hospital prison. He hacked Rockstar at a hotel using an Amazon Fire TV Stick while in police protection. Details from @joetidy https://t.co/dq3AS041Nx
Introducing SDXL Turbo: A real-time text-to-image generation model.
SDXL Turbo achieves state-of-the-art performance with a new distillation technology, enabling single-step image generation with unprecedented quality, reducing the required step count from 50 to just one.
The code, research paper, and weights for non-commercial use are now available on our website.
You can test SDXL Turbo on Stability AI’s image editing platform @Clipdropapp, with a beta demonstration of the real-time text-to-image generation capabilities.
Learn more: https://t.co/L39rZWf9F7
Today, we are releasing Stable Video Diffusion, our first foundation model for generative AI video based on the image model, @StableDiffusion. As part of this research preview, the code, weights, and research paper are now available.
Additionally, today you can sign up for our waitlist to access a new upcoming web experience featuring a Text-To-Video interface.
To access the model & sign up for our waitlist, visit our website here: https://t.co/IcuPJr45S9
Bonjour @midjourney 👋
We're Heetch, probably not on your radar, but we, along with the 12.5 million residents of the French Banlieue, have just sent out thousands of postcards that you - we mean, your 11 employees - should receive soon.
Once these postcards land on your desks, reach out! We're all ears for a chat ☺️
Best regards from La Banlieue,
Heetch 🚘🚘
PS: Feel free to check the video for more details
NEW: LLM startup @LaminiAI revealed it has been “secretly running on more than 100” @AMD Instinct MI200 series GPUs and said the chip designer’s ROCm software platform “has achieved software parity” with @nvidia's dominant CUDA platform for LLMs. https://t.co/0ZZEK2aVmH via @CRN
🌟 Introducing Falcon 180B: The World's Most Powerful Open LLM! 🚀
At #TII, we are continuing to push the boundaries of generative AI with our open access Falcon 180B AI model which has already soared to the top of the Hugging Face Leaderboard.
Code LLaMA is now on Perplexity’s LLaMa Chat!
Try asking it to write a function for you, or explain a code snippet: 🔗 https://t.co/rwcPzknBgE
This is the fastest way to try @MetaAI’s latest code-specialized LLM. With our model deployment expertise, we are able to provide you with this model less than 24 hours of it’s release.
What’s next?
We’ll integrate code LLaMA into Perplexity, all in service of providing you with the best answers to your most technical questions!
🚀Exciting news! Stability AI has launched StableCode, the revolutionary generative AI LLM for coding!
💡 Developers, get ready to level up your coding game! #AI#Coding#StableCode#StabilityAI
https://t.co/XFrV36JMMu