
The transformer architecture, which has underpinned all major large language models since 2017, is facing a fundamental scaling problem: computational demands and power consumption spike sharply as input text grows longer.
Four startups are pursuing different technical approaches to overcome this bottleneck, potentially unlocking cheaper and more energy-efficient AI systems.
What happened
The transformer architecture, which has powered every major large language model since 2017, is hitting scaling limits—the computational cost and power consumption surge as text length increases. Four startup approaches are now challenging this constraint.
Why it matters
Transformers' strength in handling sequential text has become a weakness: longer documents demand exponentially more compute and electricity. Finding alternatives could make AI systems cheaper and more energy-efficient to run, lowering the bar for smaller companies and applications.
What to watch
The article identifies four distinct technical paths; which (if any) will prove viable at scale remains open. The outcome will shape whether AI inference costs fall or remain prohibitive for resource-constrained users.
Ask the AI about this article →
The transformer, introduced in 2017, solved a critical challenge in processing sequential data: it allowed AI models to weigh the importance of different words in a long passage simultaneously, rather than processing words one at a time. This architectural innovation became so effective that it is now embedded in every major commercial large language model. However, that same mechanism—comparing every word to every other word in a document—creates a computational cost that scales poorly. As text length grows, the number of comparisons multiplies, forcing models to consume more and more electricity to produce answers. This efficiency problem has become acute as applications demand longer context windows (the amount of prior text a model can "remember" when generating answers). The article frames this as the "transformer problem," suggesting that the constraint is now inhibiting further progress. Four startups have identified this gap and are pursuing alternative directions, each betting that a different technical path can break the transformer's dominance without sacrificing the quality users have come to expect from modern LLMs.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Visko raised $10 million in pre-seed funding from Llama Ventures and opened public access to its first foundat…
AI company Runway has unveiled Solaris, the first model in a new category it calls "Interface World Models." I…

Google's AI search gave advice to call emergency services for users alone with an African, Indian, or Pakistan…

John Deere introduced JD, a conversational AI tool that lets farmers ask open-ended questions about their hist…

Nvidia CEO Jensen Huang said on Fox Business that AI is creating 'hundreds of thousands' of jobs, including in…

Israeli startup DataAgent Ltd