AIToday
Large Language ModelsOpen-Source AIAI Business & IndustryFortune AIPublished: Sep 13, 2026, 22:00 JST2 min read

Chinese AI labs close gap with cheaper 'attention' algorithms, not just alleged training on US outputs

Chinese AI labs close gap with cheaper 'attention' algorithms, not just alleged training on US outputs

3 Key Points

  1. What happened

    U.S. agencies on Tuesday accused DeepSeek, Moonshot and four other Chinese AI companies of extracting capabilities worth billions from American models since 2024. Analysts like Futurum Group's Brendan Burke point instead to algorithms that cut attention calculation complexity by an order of magnitude.

  2. Why it matters

    Hugging Face reported Chinese open-source models took 41% of downloads last year, and Ramp's index shows businesses paying for platforms with Chinese or open-source models rose to 6.1% in July from 4.5% in January.

  3. What to watch

    The shift hinges on whether cost savings outweigh the still-months-ahead performance of U.S. frontier models for sensitive tasks. Watch the $5.6 million DeepSeek training-cost figure that U.S. agencies claim is understated.

WHO IT HITSEnterprise engineering leaders and procurement teams evaluating coding tools are the clearest beneficiaries — Larridin tracks GLM 5.2 and Kimi handling around 75% of engineering tasks at a fifth of U.S. model costs. Data teams at companies like Thomson Reuters are also adapting open-source Chinese models for document review.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

The U.S.-China AI race took a sharp turn this week when the FBI, NSA and CISA alleged that DeepSeek, Moonshot and four other Chinese companies extracted 'capabilities worth billions' by training on American model outputs since 2024. China's foreign affairs ministry called the accusations 'groundless.' But the Fortune reporting suggests the allegations are only one explanation for a narrowing performance gap — a Stanford report put Anthropic's top model just 2.7% ahead of DeepSeek's earlier this year. The other explanation is structural: U.S. restrictions on Nvidia's best chips pushed Chinese labs toward domestic alternatives like Huawei, forcing them to find cheaper ways to run the 'attention' mechanism that underlies every large language model. Futurum Group's Brendan Burke describes the result as algorithms that cut computational complexity by an order of magnitude, while U.S. frontier labs, with access to 74% of the world's compute per a White House report, could afford to be 'token hogs.'

The commercial consequence is already visible in enterprise workflows. Larridin's Ameya Kanitkar says Chinese models like GLM 5.2 and Kimi 2.6 and 2.7 handle around 75% of engineering tasks at a fifth of U.S. cost, and Hugging Face reported Chinese open-source models took 41% of downloads last year. DoorDash, Cursor, Airbnb and Siemens are all cited as experimenting with or adopting Chinese models, and Ramp's index shows the share of businesses paying for platforms with access to open-source or Chinese-developed models rose to 6.1% in July from 4.5% in January. That said, the article is careful to note this is not a wholesale replacement: U.S. models remain months ahead on the most complex tasks, and AnswerRocket's Mike Finley argues Chinese labs' work 'would simply not be possible without the frontier labs blazing the trail.' The test ahead is whether cost pressure — already constraining AI use for 20% of business leaders McKinsey surveyed — pushes more enterprises toward Chinese open-weight models, or whether performance gaps on frontier tasks keep the most demanding workloads on U.S. systems.

FAQ
What technique did Chinese labs use to close the gap?
Futurum Group's Brendan Burke says Chinese labs found algorithms that reduce the complexity of attention calculations by an order of magnitude, achieving better results by summarizing the most relevant tokens.
Which Chinese models are U.S. companies actually using?
DoorDash CEO Andy Fang said Moonshot AI's Kimi was cheaper and better quality, Cursor used Kimi for its Composer 2 coding agent, and Airbnb and Siemens are experimenting with Alibaba and DeepSeek models.
Are U.S. models still ahead?
Yes. Mike Finley of AnswerRocket told Fortune U.S. AI companies' output still serves as the 'existence proof' for Chinese labs, and U.S. models remain months ahead in performance.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • GPT-6 Astra triples Claude Fable in Andon Labs testsTHE DECODER · 1h ago
  • AllSpark's Iris-mini, Iris-pro lead open-weight search agentsTHE DECODER · 1h ago
  • Yuxiang Zhou: 7 of 10 senior salespeople chose AI coach's adviceFortune AI · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleAI agents, not chatbots, drive data center buildout: Wired