
What happened
Subquadratic emerged from stealth last month claiming to have cracked a mathematical problem that has constrained large language models for nearly a decade. The company's new model, SubQ, uses sparse attention (a method that drastically cuts the number of computations needed) instead of the dense attention mechanism that powers today's most advanced models. An independent evaluation by third-party firm Appen found that SubQ was 56 times faster than models using FlashAttention, a previous sparse-attention technique, and matched the coding performance of top models from Google DeepMind, OpenAI, and Anthropic on standard benchmarks.
Why it matters
Most LLMs rely on a transformer architecture that requires multiplying every word's numerical encoding with every other word's encoding—a process that becomes exponentially more expensive as text grows longer. SubQ's sparse attention selects only the most relevant word relationships to process, which the company claims could dramatically lower costs and energy use without sacrificing performance. If SubQ's results hold up, it could reshape how companies build language models going forward and make AI applications far cheaper to operate.
What to watch
Subquadratic says SubQ can process up to 12 million tokens at once in its context window, compared with one million tokens for most top models today. The company has not yet made SubQ widely available for public testing, and cost claims are difficult to verify independently at this stage. According to the CEO, running Anthropic's Opus 4.6 through a standard test cost $2600, while SubQ cost eight dollars—but until the model is available to the broader market, this comparison cannot be independently confirmed.
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Phonely Ltd. launched Alma, a large language AI model built for voice agents and trained on over 10 million re…
Aranya Inc., a startup founded last year, launched today with $11 million in funding
CBTS Technology Solutions LLC launched Forge Agents, a platform that turns a plain-language job description in…
Imec CEO Patrick Vandenameele said at SEMICON Taiwan 2026 that the Belgian research center is broadening its c…

Alphabet's AI Overviews now reach over 2.5 billion monthly users through Google Search, and its ad business ge…

Sarah O’Connor's book 'We Are Not Machines' explores how mechanization and AI have transformed the workforce…
