6 entries on this page
Restore LLM Inference Capacity in Seconds with Shadow Engine Recovery in NVIDIA Dynamo
12d ago
When an LLM engine process fails, the standard recovery path involves a cold restart. This requires loading weights into HBM from storage, compiling kernels…
CUDA Python 1.0: Stable APIs, One Foundation, Full Platform Access
13d ago
For years, a Python developer who needed a GPU had two realistic choices: Learn NVIDIA CUDA C++ well enough to write an extension, set up a build toolchain…
Giga-Scale AI and the Ethernet Evolution: How Spectrum-X Ethernet Rewrites the Rules
14d ago
The massive growth of generative AI has fundamentally altered data center design. As distributed model training scales to span hundreds of thousands of GPUs…
NVIDIA Vera Rubin and Blackwell Set a New Standard for Agentic AI Performance per Watt
14d ago
AI agents have expanded inference from single-turn interactions into multi-step workflows that reason, invoke tools, coordinate subagents…
NVIDIA BlueField-4 Powers New Scale-In Network Infrastructure for Agentic AI Factories
14d ago
Traditional cloud infrastructure was designed for predictable, general-purpose workloads and standard interfaces. Agentic AI factories connect diverse users…
Solving Agentic AI Fleet Challenges with NVIDIA Vera CPU
14d ago