Holy crap POTUS just phoned in @JensenHuang live on stage at All In Summit
We will not lose the ai race! And whatever Dario said this weekend won’t stop our progress
This made my morning!
$NVDA
For my first post, I’m sharing a letter @NVIDIA signed on why open models matter.
AI will transform every industry, power every company, and be built by every country.
Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty.
The world needs both frontier closed models and frontier open models.
https://t.co/AUKzoQ5Ikb
1.1M tokens/sec on just one rack of GB300 GPUs in our Azure fleet.
An industry record made possible by our longstanding co-innovation with NVIDIA and expertise of running AI at production scale!
https://t.co/1qLvoS2z70
That’s for the DGX Spark 😀
This is ~100X more compute per watt than the DGX-1, the first ever dedicated AI computer, that Jensen gave me at OpenAI in 2016!
NVIDIA Blackwell can achieve 303 output tokens/s for DeepSeek R1 in FP4 precision, per our benchmarking of an Avian API endpoint
Artificial Analysis benchmarked DeepSeek R1 on an @avian_io private API endpoint. Running DeepSeek R1 in FP4 precision on NVIDIA Blackwell, their endpoint achieved 303 output tokens/s - the fastest speed we have measured yet for DeepSeek R1.
The FP4 version of DeepSeek R1 maintained accuracy across our evaluation suite as compared to the native FP8 version.
Inference speed is especially critical for reasoning models that ‘think’ before they answer - we look forward to wider availability of NVIDIA Blackwell hardware in the coming months!
Introducing DeepSeek-R1 optimizations for Blackwell, delivering 25x more revenue at 20x lower cost per token, compared with NVIDIA H100 just four weeks ago.
Fueled by TensorRT DeepSeek optimizations for our Blackwell architecture, including FP4 performance with state-of-the-art production accuracy, it scored 99.8% of FP8 on MMLU general intelligence benchmark.
FP4-optimized DeepSeek checkpoint now available on @huggingface: https://t.co/NxLukbCESw
Exclusive developer livestream on 📆 December 18th at 9 a.m. PST
Unlock high performance #generativeAI with NVIDIA NIM inference microservices - learn how to get started with this powerful tool.
Perfect for developers of all skill levels.
Register ➡️ https://t.co/gCEcRmNn2g
Join us at #GTC24 to explore the transformative journey of #cuOpt. Learn how this GPU-accelerated library has evolved into an AI cloud API, enabling recent advances in route optimization. Register today: https://t.co/vqeGte6kBy
My team at NVIDIA is hiring. We 🩷 you all from OpenAI. Engineers, researchers, product team, alike. Email me at [email protected]. DM is open too. NVIDIA has warm GPUs for you on a cold winter night like this, fresh out of the oven.🩷
I do research on AI agents. Gaming+AI, robotics, multimodal LLMs, open-ended simulations, etc. If you want an excuse to play games like Minecraft at work - I'm your guy.
I'm shocked by the ongoing development. I can only begin to grasp the depth of what you must be going through. Please, don't hesitate to ping me if there's anything I can do to help, or just say hi and share anything you'd like to talk about. I'm a good listener.