Fetching from the wire…
Infra2026-08-27 · source-backed
NVHBM, announced August 26, relocates NVIDIA's custom memory controller from the compute chip into the HBM base die, claiming up to 30% more bandwidth than standard HBM4E, 15% lower HBM power, and up to 25% more freed area on the XPU compute die (NVIDIA). Next-generation Trainium4 will support NVHBM alongside NVLink Fusion. Nvidia selling memory architecture into the custom silicon that competes with its GPUs is the strategic move, and it's the same week AWS and Nvidia announced 2 million additional GPUs deployed across AWS infrastructure over two years (GlobeNewswire), while Amazon AI chief Peter DeSantis has said AWS is in talks to sell Trainium externally as a direct Nvidia alternative.
Each link below shares sources, entities, or timing with this story.
The Hrazdan facility opened August 8, scaling to 300 megawatts and 70,000+ NVIDIA Rubin and Blackwell GPUs by end of 2027, built on NVIDIA DSX (40% more GPUs on the same footprint) with Dell PowerEdge, Schneider Electric power and Vertiv cooling. NVIDIA intends to invest, foll...
TechCrunch toured Amazon's Trainium lab. Trainium2 is now a multi-billion dollar business growing 150% QoQ with 1.4M chips deployed. Anthropic runs Claude on over 1 million. Apple is testing Trainium. AWS custom silicon is the first structural threat to NVIDIA's near-monopoly....
Announced at FMS, the cuFile APIs let GPUs read and write storage directly in microseconds rather than routing through CPUs, and SCADA (scaled, accelerated data access) lets massively parallel GPUs pull only application-necessary data into high-bandwidth memory. NVIDIA also sh...
TechCrunch's August 29 piece frames Nvidia's durable advantage as system-level, built around Vera Rubin pairing the Rubin GPU with the Vera CPU, a Groq 3 LPX inference accelerator, and storage and networking racks. VP of storage technology Jason Hardy is quoted claiming "upwar...
NVIDIA claims 1.8x faster task completion and twice the efficiency against traditional x86, with Vera Rubin NVL72 racking 72 Rubin GPUs and 36 Vera CPUs alongside ConnectX-9 SuperNICs and BlueField-4 DPUs. (NVIDIA) Architecture detail months before shipping, two days ahead of...
Announced August 31, with MediaTek adopting NVLink Fusion so custom accelerators built for hyperscalers still plug into Nvidia rack-scale systems, plus extensions to the existing RTX Spark and DGX Spark PC chip work and MediaTek's Dimensity Auto line (TechCrunch). Nvidia is co...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.