Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles.
Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts.
It’s a powerful backbone for robotaxis, trucks, shuttles, delivery vans, tractors and the long tail of mobile robots—billions of autonomous machines someday.
We’re releasing it for commercial use under OpenMDW-1.1 so teams can inspect it, fine-tune it and deploy it—open models advance safety and security.
The next wave of AI is robotics—and it starts with autonomous vehicles.
Great work, Alpamayo team!
https://t.co/2PYCCXWjZh
Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community.
During the Hugging Face incident, closed AI blocked essential forensics. An open-weight frontier model helped contain the intrusion.
That’s why we created the Open Secure AI Alliance.
For my first post, I’m sharing a letter @NVIDIA signed on why open models matter.
AI will transform every industry, power every company, and be built by every country.
Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty.
The world needs both frontier closed models and frontier open models.
https://t.co/AUKzoQ5Ikb
Thank you to the partners across America helping bring the supply chain home.
In 43 states and growing, NVIDIA’s network of American partners and suppliers spans semiconductors, boards, systems, racks, and more.
Together, we’re helping create the tools for better healthcare, breakthrough scientific discovery, stronger industrial productivity, and global technology leadership.
Water usage has been a hot topic in the AI data center world, but the numbers may surprise you.
According to the Manhattan Institute, data centers use 0.2 percent of daily water usage in the U.S. and that number has dramatically decreased in the past few years due to a new method: liquid cooling.
By moving to 45°C liquid cooling, AI factories in favorable climates can use dry coolers instead of conventional cooling-tower-based systems, cutting facility cooling water use from roughly 2.6M gallons per MW per year to near zero.
Liquid cooling enables AI factories to be both water and energy efficient, while creating opportunities for heat reuse and dispersal to local communities, allowing these factories to become energy grid assets.
Learn more below ⬇️
https://t.co/7WanoPNKTR
Amazing work from the @sgl_project and @radixark team for their work optimizing DeepSeek V4 inference on B200, B300, and the recent 4x iso-interactivity throughput improvements on GB300 by @ChengWan17! As @elonmusk said, The GB300 is the best AI computer, and software optimizations like this show its true potential!
Internally at NVIDIA, we use cuOpt based agentic workflows with agent skills to optimize our supply chains. Since it’s open source, you can too.
With optimizations ready in minutes instead of weeks, the workflow uses multi-agent LangChain Deep agent orchestration and GPU-accelerated solvers to turn natural language into optimized decisions.
Spin it up instantly with a Brev Launchable (preconfigured GPU environment) and grab free developer credits while they last.
MINECRAFT STEVE ALERT: GB300 ultra NVL72 is already 2.7x faster 🚀 than GB200 NVL72 on one of the industry standard inference engine known as @vllm_project. On paper, GB300 only has ~1.5x faster NVFP4 FLOP & 1.5x more HBM capacity & same HBM BW than GB200 but due to the full stack optimization with compounding gains, in the middle of the curve where most providers serve at, GB300 is up to 2.7x faster. End to End performance is the gold standard of performance, not on paper theoretical flops.
Thanks to the 10x engineers at NVIDIA & @inferact & @coreweave for this temporary gb300 for open source projects!
✨ DeepSeek-V4 is here — a million-token context, 1.6T parameter powerhouse optimized for agentic workflows.
Out of the box, on DeepSeek-V4-Pro, NVIDIA Blackwell Ultra delivers over 150 TPS/user interactivity for agentic workflows.
And we’re just getting started. Expect these performance figures to climb higher as we implement Dynamo, NVFP4, and advanced parallelization techniques.
Start building today with @lmsysorg and @vllm_project
At GTC 2024, Jensen said that GB200 NVL72 was 35x faster than Hopper. Nobody believed it and thought it was classic fake Jensen Math. When we tested the performance of it, it wasn't just 35x faster, it was over 50x times faster even against an strong Hopper baseline with all of the inference optimization composed together like MTP, Disagg prefill, wideEP, etc. View the nuanced results at InferenceX dot com.
NVIDIA just killed the awkward pause in voice AI 😱
PersonaPlex 7B is a real-time conversational model that listens AND speaks simultaneously. Like actually interrupts you mid-sentence like a human.
Beat Gemini Live on dialog naturalness. 18x faster interruptions.
100% open source. Run it locally. No API bill. No latency.
💡 @jumptrading is accelerating AI-powered research and financial modeling as one of the first trading firms to adopt the NVIDIA Rubin platform and Vera Rubin NVL72—delivering supercomputer-class compute density and enhanced performance.
Learn more about the future of algorithmic trading: https://t.co/RcmioelMhp
#NVIDIAGTC
NVIDIA cuOpt is officially the #1 OSS solver on the Hans Mittelmann MIPfeas leaderboard. This is a massive win for GPU-accelerated Mixed Integer Programming, proving cuOpt is ready to power the next generation of complex, memory-intensive workloads.
From complex fleet routing to supply chain planning, the possibilities for low-latency agentic systems just expanded.
Hans Mittelmann Benchmark: https://t.co/sx5OaBTL8P