
The AI chip market is diversifying beyond NVIDIA's GPUs.
CPUs and LPUs now serve agentic AI tasks.
Edge AI chips from Japan and the US bring generative AI offline.
What happened
AI accelerators that once centered on GPUs now include CPUs, LPUs, and more. Companies like AMD, Google, and startups such as Cerebras and Groq are developing cloud AI chips, while EdgeCortix and SiMa.ai target edge AI.
Why it matters
As AI evolves from deep learning to generative and agentic AI, chips must handle varied tasks. CPUs suit agentic workflows, and LPUs like Groq's speed up token generation. NVIDIA's CUDA software strength keeps it leading in semiconductor sales.
What to watch
AMD's Helios with 6thGen EPYC and MI400 GPUs, paired with Cerebras's Wafer Scale Engine 3, achieved 5x faster token generation speed (Tokens/second/W). EdgeCortix's SAKURA-II offers 60TOPS at 10W, and SiMa.ai's Modalix device costs about 20万円 (Japanese yen).
Ask the AI about this article →
Summaries like this, in your inbox every morning.
The article traces AI chip evolution from 2012, when Geoffrey Hinton's lab showed GPUs cut image recognition errors, a moment NVIDIA's CEO calls the 'big bang of AI.' That breakthrough led to deep learning and later generative AI, but training took hundreds of days even with thousands of GPUs. This demand for performance drove innovation beyond NVIDIA, yet NVIDIA's lead persists partly because its CUDA software ecosystem makes its chips easier to use. Now, with agentic AI, the article argues CPUs are better suited for workflow-driven tasks, and with physical AI, edge devices need on-device generative capabilities. This diversity means no single chip type dominates. However, designing advanced AI chips at 2nm is so complex that only a few design houses may handle it. That could limit which companies can produce data-center AI chips, potentially benefiting firms like Broadcom, Marvell, MediaTek, and Japan's Socionext, possibly aiding efforts to revive Japan's semiconductor industry.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
JM-Applied, a Taiwanese semiconductor gas equipment maker and supply chain member for Micron and TSMC, said on…

OpenAI's product lead Tibo Sotiou posted on X on September 6 that GPT-6 Astra's 'low' setting outperforms GPT-…

Nvidia disclosed roughly $99 billion of public and private equity investments as of July 26, plus about $25 bi…

In January, Ukraine's defense ministry said it would share millions of data points from tens of thousands of d…

OpenAI Group PBC acknowledged it did not publicly disclose an episode where its AI agents wrote to outside web…
Eaton is expanding beyond traditional power management into modular power deployment, next-generation DC conve…
