
AI's memory needs are expanding beyond HBM into DRAM and storage.
Bernstein says the investment opportunity depends on the AI workload.
KV caches could limit how many users an AI service can support.
What happened
Bernstein analysts say AI is creating a memory bottleneck, with demand expanding from high-bandwidth memory (HBM) into conventional DRAM, NAND flash, and storage. They rate Samsung Electronics, SK hynix, Micron, SanDisk, Seagate, and Western Digital as Outperform, while Kioxia is rated Underperform.
Why it matters
Different AI workloads—training, inference, retrieval-augmented generation, and agentic AI—place different demands on memory. In inference, the "decode" stage is memory-bound and relies on a key-value (KV) cache, which could require more memory than model weights in large deployments, potentially limiting how many users an AI service can support.
What to watch
Emerging memory tiers, including CXL memory, Nvidia's "Storage Next" initiative, and CMX context storage, aim to balance performance, capacity, and cost. Bernstein notes that technical hurdles remain high for high-bandwidth flash, which seeks to combine HBM-like bandwidth with NAND's greater capacity and lower cost.
Ask the AI about this article →
The article highlights a shift in AI's memory demands, moving beyond the initial focus on high-bandwidth memory (HBM) to include conventional DRAM, NAND flash, and storage. This is driven by the different requirements of AI workloads—training, inference, retrieval-augmented generation, and agentic AI—each placing unique pressures on the memory stack. For instance, while training is compute-intensive and relies heavily on HBM, the decode stage of inference is memory-bound, with KV cache sizes growing with context length and concurrent users. This could limit service capacity, a key concern for AI providers.
The emergence of new memory tiers, such as CXL memory and Nvidia's "Storage Next," reflects an industry effort to manage performance, capacity, and cost. However, technical challenges remain, particularly for high-bandwidth flash, which aims to bridge the gap between HBM's bandwidth and NAND's capacity. Bernstein's ratings suggest confidence in established memory manufacturers, but the dynamic nature of AI workloads means investment opportunities will depend on how these technologies evolve.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
OpenAI announced on August 28 that it will terminate its model access contract with AI coding tool Cursor, pro…

OpenAI CEO Sam Altman said in a Time magazine interview that he thinks "it is a good time to slow down" AI mod…

Marvell Technology reported record fiscal Q2 2027 revenue of $2.739 billion, up 37% year over year, and raised…

NVIDIA reported $96.2 billion in revenue for fiscal Q2 2027, up 106% year over year, with Data Center revenue…

Intel expanded its partnership with Kasm Technologies to support compliant, local AI workloads on Intel Xeon 6…

In the first quarter of fiscal 2027, Dell booked $24.4 billion in AI orders and generated $16.1 billion in AI-…
