Nvidia unveils NVHBM memory architecture, moving controller into HBM stack to free 25% more compute
Nvidia announced NVHBM, a new memory architecture that relocates the HBM memory controller from the AI accelerator die into the memory stack itself, freeing up to 25% more compute silicon per chip. The move addresses AI infrastructure bottlenecks as trillion-parameter models become mainstream, while Nvidia expands its NVLink licensing strategy amid growing custom chip competition.