NVIDIA Unveils NVHBM as AWS Lines Up Trainium4 Support and a Major GPU Deal
NVIDIA has introduced NVHBM, a new custom high-bandwidth memory technology, and folded it into its NVLink Fusion ecosystem as AI infrastructure demands shift beyond raw compute toward tighter coordination across processors, memory, interconnects, and cooling. AWS chip unit Annapurna Labs is the first partner to adopt the technology and said its next-generation in-house Trainium4 chip will support both NVLink Fusion and NVHBM. The announcement came alongside a separate commercial commitment: Amazon confirmed plans to buy up to 2 million next-generation high-end GPUs from NVIDIA between 2027 and 2028, covering Blackwell Ultra, Rubin, and Rubin Ultra. According to ABMedia, the arrangement reflects a changing relationship between hyperscalers and NVIDIA, where internal chip development and external GPU procurement now sit side by side rather than in direct opposition. ABMedia said NVHBM moves the memory controller into the base die of a 3D HBM stack, freeing die space on the XPU for more compute cores or cache. The report cited gains of up to 30% more memory bandwidth, about 15% lower HBM power consumption, and up to 25% more chip area released for compute. NVIDIA is also pushing the design toward a standardized specification across suppliers including SK hynix, Samsung, and Micron.



