NVIDIA NVLink Fusion Expands With NVHBM Custom High-Bandwidth Memory
The next wave of AI is placing new demands on infrastructure. As AI agents and trillion-parameter workloads become mainstream, the performance of AI infrastructure depends not only on compute, but on how compute, memory, storage, networking and software are designed together as a

Amazon’s Annapurna Labs will be the first to collaborate on NVHBM technology alongside NVLink Fusion.
August 26, 2026 by Jesse Clayton
0 Comments
-
X
-
-
- [
Copy link
Link copied!
](#)
The next wave of AI is placing new demands on infrastructure.
As AI agents and trillion-parameter workloads become mainstream, the performance of AI infrastructure depends not only on compute, but on how compute, memory, storage, networking and software are designed together as a unified system.
To help hyperscalers and AI innovators build the next generation of semi-custom AI infrastructure, NVIDIA today expanded NVIDIA NVLink Fusion with NVIDIA NVHBM, a next-generation high-bandwidth memory technology that brings higher memory performance and efficiency to XPUs. It will be validated and offered by leading memory partners, extending this advanced memory capability to NVLink Fusion customers.
Traditional HBM architectures place the memory controller on the XPU die, consuming valuable silicon area that could otherwise be dedicated to compute. NVHBM, built on the same technology that NVIDIA will use for future GPUs, integrates NVIDIA’s custom memory controller into the HBM base die.
By integrating the memory controller into the 3D HBM stack instead of the XPU, NVHBM delivers up to 30% greater memory bandwidth and 15% lower HBM power consumption, and frees up to 25% more area on XPU compute die compared with standard HBM4E.
NVIDIA is establishing a standard NVHBM implementation, available from multiple memory providers. This reduces the engineering effort required to integrate and qualify memory across multiple suppliers — giving NVLink Fusion customers a faster path for bringing custom AI chips to market.
Amazon’s Annapurna Labs will be the first to work on NVHBM as part of its broader collaboration with NVIDIA around NVLink Fusion.
AWS and NVIDIA Continue NVLink Fusion Collaboration
Amazon’s Annapurna Labs will work with NVIDIA on NVHBM technology and the NVLink scale-up architecture to enhance performance and efficiency for AI workloads.
This builds on AWS’s previously announced support for NVLink Fusion. Annapurna Labs will support NVLink Fusion with its next-generation Trainium chips starting with Trainium4, which would allow Amazon chips and NVIDIA GPUs to work together with common rack-scale architecture.
“NVHBM represents a new architectural approach to advancing high-bandwidth memory performance and efficiency,” said Nafea Bshara, vice president of Annapurna Labs at Amazon. “We look forward to this technology collaboration to benefit future AWS infrastructure designs.”
Vertically Integrated and Horizontally Open
NVLink Fusion enables partners to connect custom XPUs and CPUs to NVIDIA’s rack-scale platform.
Partners can access NVIDIA NVLink chiplets, NVLink-C2C, NVLink Switches and NVIDIA MGX systems and racks, as well as a broad ecosystem of CPU partners, ASIC designers, system manufacturers and technology providers.
Offered with each generation of NVIDIA’s rack-scale system architecture, NVLink Fusion allows hyperscalers and AI-native companies to focus engineering resources on XPU innovation while using a proven technology stack for scale-up and scale-out networking, rack-scale systems and software — creating a faster, lower-risk path to deploying semi-custom AI infrastructure.
Learn more about NVLink and NVLink Fusion.
- Categories:
- AI Infrastructure
- Hardware
- Networking
- Tags:
- NVLink
Related News
AI Infrastructure
Delivering Vera: NVIDIA’s First CPU Built for Agents Is Shipping Now
Aug 27, 2026
AI Infrastructure
NVIDIA NVLink Fusion Expands With NVHBM Custom High-Bandwidth Memory
Aug 26, 2026
AI Infrastructure
With Groq 3 LPX in Full Production, NVIDIA Extends Vera Rubin Inference for Agents
Aug 24, 2026
AI Infrastructure
Up to 30x More Work Per Watt: NVIDIA Vera Rubin NVL72 Sets a New Efficiency Standard for AI Agents
Aug 24, 2026
Related stories
Samsung Galaxy S26 FE: Delivering the Latest Flagship Experience, Focused on What Matters Most
Samsung Newsroom
Samsung Electronics today announced Galaxy S26 FE, the newest addition to the Galaxy S26 family and the first in the lineup to launch with One UI 9 — bringing the latest premium Galaxy experiences to more users from day one. With enhanced camera capabilities and more context-awar

How AI Coding Agents Can Unlock Materials Simulation with NVIDIA ALCHEMI Toolkit
NVIDIA Developer Blog
Atomistic simulation requires three things: knowledge of the science, compute-efficient implementation of simulations, and accessible interfaces to the... Atomistic simulation requires three things: knowledge of the science, compute-efficient implementation of simulations, and ac
Samsung Introduces New Odyssey Lineup for Fast-Paced Gaming at Gamescom 2026
Samsung Newsroom
Samsung Electronics today announced its 2027 Odyssey gaming monitor lineup at Gamescom 2026, the world’s largest gaming event, being held in Cologne, Germany from Aug. 26-30. The new lineup introduces multiple Odyssey models that feature world-first innovations, empowering player

How we saved 100 terabytes of memory by optimizing 1.1.1.1’s DNS cache
Sebastiaan Neuteboom
Big Pineapple , the platform behind 1.1.1.1 , Gateway DNS , DNS Firewall , AS112 , and several other Cloudflare DNS services, stores over 250 billion DNS cache entries at any given time. At that scale, wasting a single byte per entry costs more than 250 gigabytes of memory across

The Most Efficient Token Is One You Don’t Spend
Dell Blog
Why better enterprise data is the foundation of a more affordable AI future.
Scott Pilgrim EX: Bringing three new heroes to life in Back in the Band DLC
Eric Lafontaine
When we decided to reunite the full Sex Bob-omb lineup for Scott Pilgrim EX – Back in the Band, available now, we knew that simply adding three new playable characters wasn’t enough. Stephen Stills, Kim Pine, and Knives Chau each needed to feel like they had always belonged in th
