
Nvidia is testing its next Rubin GPU with less memory than promised as an industrywide shortage intensifies.
AMD claims it has already secured sufficient supply for its competing Helios system shipping this year.
The memory cut could force AI companies to buy more Nvidia chips per model, shifting the cost equation.
What happened
Nvidia is testing Rubin Ultra GPU versions with 192GB to 256GB of high-bandwidth memory, well below the 1 terabyte Jensen Huang originally announced, according to The Information. Meanwhile, AMD is shipping its Helios system to Microsoft, Meta, OpenAI, and Oracle later this year, claiming it has already secured the memory it needs through relationships with Samsung Electronics, SK Hynix, and Micron Technology.
Why it matters
The memory downgrade raises questions about whether Nvidia is genuinely caught off guard by industry-wide memory shortages, or whether AMD's confidence in locking up supply is the bigger story. Running less memory per GPU could force AI companies to buy more chips to run the same models, adding cost even if unit prices drop—a reversal from Nvidia's July claim that it was "in front of the memory problem."
What to watch
Rubin Ultra does not ship until late 2027, giving Nvidia time to adjust. Nvidia has also struck a $500 billion partnership with SK Hynix's parent company to co-develop future memory technology. AMD reports that data center revenue is its fastest-growing segment and, as of Q1 2026, makes up the majority of AMD's total revenue, up 57% year over year, with Meta planning to deploy up to six gigawatts of AMD chips.
Ask the AI about this article →
The memory shortage facing the AI chip industry is forcing a reckoning between the two dominant players. Nvidia, which controls more than 95% of the data center GPU market, had publicly projected confidence just weeks before—hardware SVP Andrew Bell said in mid-July that the company was "in front of the memory problem." Yet testing of reduced-memory Rubin variants suggests either that confidence was misplaced or that Nvidia is preemptively adapting to a constrained supply reality. The versions under test (192GB to 256GB, versus the 1 terabyte originally announced) would represent a significant cut.
AMD's rival positioning is notably different: the company claims it has already locked up high-bandwidth memory allocations through direct relationships with all three major suppliers—Samsung Electronics, SK Hynix, and Micron Technology—and has secured major customer commitments from Microsoft, Meta, OpenAI, and Oracle. AMD's data center revenue was up 57% year over year as of Q1 2026 and now constitutes the majority of the company's total revenue, signaling real momentum. Whether AMD's supply confidence will hold and whether Helios can genuinely compete on workloads where Nvidia dominates remain open questions, but the divergence in their public posture reflects a genuine shift in competitive leverage around memory access.
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
Ask AI anything about this article. Q&As are published on this page for other readers too.
ABC News will air an exclusive interview with Texas Governor Greg Abbott on Sunday, August 23, 2026, on the pr…

Western Digital has returned 530.0% over the past year and 14.6x over three years, yet trades on a P/E multipl…

Palo Alto Networks' Unit 42 expanded its Frontier AI Exposure Analysis service by integrating Anthropic's Clau…

SpaceXAI released Grok 4.6—its latest flagship model—on Google Cloud's Vertex AI platform on August 21, 2026…

Eric Schadt, a computational biology expert and former chief scientist at Pathos AI, is departing that role to…

Blackstone, NVIDIA, and several global financial institutions have signed memorandums of understanding to deve…