
A team from UMD, UVA, WUSTL, UNC, Google, and Meta created AutoTTS, which uses Claude Code to search for better control algorithms in a simulated environment instead of having humans write them by hand. The discovery run cost about $40 and took 160 minutes.
The agent-discovered algorithm tracks how a model's confidence shifts over multiple rounds and adjusts the number of solution paths (width) and their depth dynamically—logic the authors call something humans probably wouldn't have designed by hand. On math benchmarks like AIME and HMMT, it achieves better accuracy per unit of compute than established methods, slashing token usage by about 70 percent compared to standard self-consistency while holding accuracy steady.
The algorithm transfers to different models (tested on DeepSeek-R1-Distill-Llama-8B) and different benchmark types (GPQA-Diamond), suggesting the approach generalizes beyond the initial discovery environment.
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
CBTS Technology Solutions LLC launched Forge Agents, a platform that turns a plain-language job description in…
Imec CEO Patrick Vandenameele said at SEMICON Taiwan 2026 that the Belgian research center is broadening its c…

Alphabet's AI Overviews now reach over 2.5 billion monthly users through Google Search, and its ad business ge…

Amazon Web Services (AWS) has integrated its fully managed data warehouse service, Amazon Redshift, with Agent…

Visual Studio Code 1.135 now includes an experimental 'Rubber Duck' feature that lets developers request a sec…

Anthropic is making a permanent 25% increase to the usage limits in Claude Code, effective after September 14
