AI Briefing
KO

NVIDIA Expands NVLink Fusion with NVHBM Custom High-Bandwidth Memory

·2026.08.27 06:05

Key point

NVIDIA added NVHBM to NVLink Fusion to enhance XPU memory bandwidth and efficiency.

Details

The mainstreaming of AI agents and trillion-parameter workloads is presenting new requirements for infrastructure. NVIDIA has expanded NVLink Fusion with NVHBM (NVIDIA NVHBM) technology to enable hyperscalers and AI innovators to build next-generation semi-custom AI infrastructure.

Existing HBM architectures placed memory controllers on the XPU die, consuming silicon area that could otherwise be used for computation. NVHBM addresses this by integrating NVIDIA's custom memory controllers into the HBM base die. This delivers performance benefits including 30% higher memory bandwidth, 15% lower power consumption, and 25% more XPU compute die area compared to standard HBM4E.

NVIDIA has established a standard NVHBM implementation available from multiple memory suppliers, reducing engineering burdens for customers and accelerating the time-to-market for custom AI chips. Amazon's Annapurna Labs is leading the development of NVHBM technology as part of the NVLink Fusion collaboration, with support planned for the next-generation Trainium4 chip. This enables Amazon chips and NVIDIA GPUs to operate together within a common rack-scale architecture.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.