AIToday
Large Language ModelsAI Coding AssistantsAINOWPublished: Sep 25, 2026, 22:00 JST

Claude Code, Devin reshape coding as AI agents take over

Claude Code, Devin reshape coding as AI agents take over

3 Key Points

  1. What happened

    AINOW's guide says AI coding agents now run planning, implementation, testing and fixes on their own, citing a Stack Overflow survey where about 31% of developers already use them at work.

  2. Why it matters

    The gap is widening between teams that delegate routine implementation and those that wait, the article argues, so design and review time can replace hands-on coding.

  3. What to watch

    Gains reported by Rakuten and Nubank are single-company cases, so the test is whether your team can start from small, well-defined tasks and grow that scope safely.

WHO IT HITSEngineering managers and developers evaluating coding tools are the primary audience, since the article's suggested entry point is low-risk work like adding tests or fixing minor bugs. Security, legal and finance teams also have a stake, given the article's warnings on leaked credentials, license-tainted code and usage-based billing.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The article frames AI coding agents as a distinct step beyond autocomplete: rather than predicting the next line, an agent keeps a work plan, edits files, runs tests and revises until the task passes. It sorts the tools into three styles — terminal agents like Claude Code that work across a whole repository, IDE agents like Cursor that show diffs for approval, and cloud agents like Devin that take a task and return a pull request. Rule files such as CLAUDE.md and AGENTS.md carry project conventions between sessions, and permission settings decide how much the agent may do without asking.

The evidence the article offers is mostly company case studies. Rakuten had Claude Code work for 7 hours on vLLM, a 12.5-million-line open-source library, and reported cutting lead time from 24 business days to 5. Nubank used Devin on a migration of more than 6 million lines, reporting that work once forecast at 18 months and over 1,000 engineers finished in weeks at more than 20x lower cost. Classmethod reported up to 10x productivity and an 80% cut in code-review time, and ULS Group used Devin on about 1.5 million test steps, cutting effort from 200 person-months to 50. Set against this, Veracode's 2025 survey found 45% of code samples from over 100 AI models failed security tests, and 46% of developers told Stack Overflow they do not trust AI tool output.

That contrast is the crux. The outcome likely hinges on process rather than tool choice — clear task boundaries, limited permissions, rule files and automated test and CI checks. Teams with those in place may be able to widen what they delegate; teams without them risk carrying generated errors into production. The article's own answer to the skeptics is to start small, on tasks where success or failure is easy to judge.

FAQ
What are AI coding agents and how do they differ from code-completion tools?
Code-completion tools suggest the next line while a person drives, with the open file as their main scope. AI coding agents take a whole task — adding a feature or fixing a bug — plan it, edit files, run tests and open a pull request, with the person reviewing the result.
What results have companies reported from using these agents?
Rakuten said Claude Code worked autonomously for 7 hours on the open-source vLLM library and cut new-feature lead time from 24 business days to 5, a 79% reduction. Nubank reported completing some migrations in weeks instead of months and cut costs by more than 20x.
What should teams check before rolling out an AI coding agent?
Decide which tasks to hand over, set permission limits on file edits and command execution, and add project rule files such as CLAUDE.md or AGENTS.md. Also guard against credential leaks, unintended commands, license violations and swelling usage-based costs.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Anthropic signs $11.6 billion Akamai cloud dealTHE DECODER · 1h ago
  • Google tests "Call for Me" on Pixel 11THE DECODER · 1h ago
  • Microsoft's Copilot super app targets business usersFortune AI · 1h ago

AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleApple Intelligence for Home: $60 vs $20, Ring's Unusual Event wins