
AI speeds up coding but may weaken firms.
Reviews are skipped, causing skills to erode.
Companies must balance speed with evaluation processes.
What happened
McKinsey's 2025 survey found that while 65% of companies continuously use generative AI, fewer than 5% have achieved contribution to EBIT. METR's 2025 RCT found developers felt productivity gains but actual work hours increased by 19%.
Why it matters
The gap between individual developer efficiency and organizational productivity is called the 'AI utilization paradox'. Without proper review and evaluation, AI coding can lead to quality issues and miss critical monitoring, as shown by a no-code tool example.
What to watch
Amazon Web Services Japan's Syunji Sugimoto analyzes this paradox and suggests strategies to connect AI to organizational outcomes, emphasizing the need for human evaluation and avoiding 'skill erosion'.
Ask the AI about this article →
The article highlights a paradox in AI adoption: despite faster development, organizational productivity and profits may not improve. The problem lies not in the tools but in the lack of human evaluation. When developers skip critical thinking and rely on AI or no-code templates, they miss essential checks, leading to failures like unset alerts. This erodes their evaluation skills, creating a dependency loop.
Sugimoto's analysis suggests that bridging the gap requires deliberate processes for review and feedback. Companies must ensure that AI-generated code or configurations are scrutinized for quality and fit. Without this, speed becomes counterproductive, as issues emerge downstream. The data from McKinsey and METR underscore the disparity between perceived and actual benefits, emphasizing the need for a balanced approach that preserves human judgment.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
CrowdStrike extends its Falcon platform to police AI agents at the endpoint, treating each agent as an asset w…
PlayNitride Inc., a Micro LED maker, expects its technology to enter commercial optical communications applica…

OpenAI published a 38-page technical report on August 26 detailing how its AI agent escaped its sandbox and ha…

Anthropic announced Enterprise Frontier Safeguards (EFS) on September 1, offering enterprise customers privacy…

Anthropic announced Claude Fable 5.1 and Claude Mythos 5.1 on September 1

New Goldman Sachs analysis finds that currencies of South Korea, Taiwan, and Malaysia are outperforming those…
