
What happened
Deepseek released open-source programming tools for Huawei's Ascend chips, including libraries for computation and data movement, and the firms optimized a supernode of 128 Ascend 950 chips. Huawei said it "fully supported" the work.
Why it matters
Programming tools are what let outside developers write software for a chip, so open-sourcing them makes it easier for others to build on the 128-chip supernode rather than on a closed toolkit.
What to watch
Whether outside developers adopt TileLang is the test, since the release only supplies the tools, not proof that they win users. Watch Huawei's pledge that its new AI processors and supernode systems will be widely used for model training next year.
WHO IT HITSChip and model developers inside China who write code for domestic AI accelerators are the ones this lands on: open tools lower the barrier to building for Ascend hardware. The same release may worry teams outside China that had treated software lock-in as Nvidia's durable defense.
Summaries like this, in your inbox every morning.
Deepseek has been using TileLang for roughly a year and first tested it on older Nvidia chips, so the language is not new to the company. What is new is putting it behind Huawei's Ascend hardware as an open release, assembled with Huawei's support and aimed at a cluster of 128 Ascend 950 chips. The stated logic is that an independent software ecosystem for AI chips needs a universal language that is easy to program but still extracts full hardware performance; Deepseek argues TileLang fits that role better than CUDA.
The context the article supplies is that software, not silicon, has been the gap. Nvidia's position rests partly on an estimated four million CUDA developers worldwide, an ecosystem rivals like AMD have not crossed even when their hardware matched on paper. Chinese model makers such as Z.ai and Moonshot AI have moved faster than the country's chipmakers, and Huawei wants to close that gap: two weeks before Deepseek's announcement it unveiled new AI processors and supernode systems, promising wide use for model training next year. Huawei also says it cannot meet domestic demand and plans to sell fewer chips abroad, with rotating chairman Eric Xu framing the company's stance around US export controls.
Whether this release changes much depends on developers outside the two companies choosing TileLang, and on how quickly Huawei's hardware is actually used for training. The rival reading in the article is that Nvidia's advantage may be shifting rather than ending: SemiAnalysis judged the CUDA moat "potentially dead" after Jalapeño, yet still found Nvidia ahead on the harder agent workloads, and noted that Huawei's CANN stack was the only one besides CUDA to support DeepSeek V4 on day one.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
Gabelli Dividend Growth Fund started a 1.2% position in Microsoft Corporation (NASDAQ:MSFT) near its 52-week l…

Rosenblatt raised its Amazon price target to $360 from $335, kept a Buy rating, and called concerns that AI sh…

In Oracle's 80-turn evaluation, Oracle AI Agent Memory held input near 1,300 tokens per request while flat his…

Arrowfly launched AI for Engineers, combining a news desk, a Sept

LinkedIn CEO Dan Shapero told the Wall Street Journal that job seekers have sent out 30% more applications tha…

Confluent's 2026 Data Streaming Report found just 17% of Japanese firms run agentic AI in production, the lowe…
