
OpenAI's official case study says Asana's browser agent test cut cost to 76 times lower and became 5 times faster, averaging $0.47 and about 4 minutes per run across 144 tests with 4 models.
Summaries like this, in your inbox every morning.
The Asana test is presented as an official OpenAI case study, not an independent benchmark, and the numbers are explicitly described as test values with confidence rated medium. That framing matters because the mechanism behind the saving is procedural: the agent had been resending the record of pages it had already read at full price on every call, and the fix was to change how that record is kept and to batch-and-discard screenshots. Model choice was revisited as well.
The companion data point in the same roundup is LegalOn, which reported cutting Codex costs by 65% by routing different tasks to different models. Together the two examples point at a pattern where cost control comes from matching model and context handling to the task, rather than from a single stronger model.
A separate figure in the roundup gives a sense of scale on the research side: work that would take one to two months by hand was completed in about a week by letting the AI run experiments, across 144 tests with 4 models. The cost angle connects to the broader theme of the day — that how AI is used and governed, more than the AI itself, is where results diverge.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
In an Arm vs. Marvell Technology comparison, the analyst picked Marvell, pointing to its 65.2x forward P/E ver…

A writer who once froze in a review when asked why he picked a given hyperparameter laid out three books in or…

Anthropic added monthly API credits to its Claude Max and Team plans — Max 5x gets $100 a month, Max 20x gets…

Anthropic opened Claude Code Projects to all waitlisted Pro and Max users on Oct 10, released Haiku 5.5 on Oct…

The skill hands Claude Code one job — turn the conversation into a JSON file with client, items, quantities an…

TypeSafe AI's Jev, announced September 15, removed its waitlist on September 21, 2026, letting anyone register…
