AIToday
Large Language ModelsAI Coding AssistantsZenn AI/MLPublished: Oct 6, 2026, 22:00 JST

Claude Code agent teams pass tests without reading code

Claude Code agent teams pass tests without reading code

3 Key Points

  1. What happened

    In Claude Code, enabling the experimental agent teams feature (via the environment variable CLAUDE_CODE_EXPERIMENTAL_AGENT_TEAMS=1) let a builder agent and a test agent talk directly. The test agent wrote tests for the splitBill function without ever reading the implementation, using only specs the builder provided — and the tests passed.

  2. Why it matters

    The experiment suggests that separating who builds from who tests can produce passing tests without the test writer seeing the code, which may reduce the chance of tests being shaped by the implementation. Agent teams is experimental and off by default, though, so this is not standard behavior.

  3. What to watch

    Teams consume significantly more tokens than a single conversation and add coordination overhead, so the official guidance is to try subagents first and use teams only when agents need to consult each other. Watch whether agent teams moves out of experimental status.

WHO IT HITSDevelopers and engineering teams using Claude Code to automate coding tasks — especially those weighing whether to delegate test-writing to a separate AI agent — will need to factor in the heavier token cost of teams versus subagents.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The article is a practical guide for people who already know Claude Code basics — asking, permission modes, CLAUDE.md, /clear and /rewind — and want to hand over more work. It groups five features from videos #1–#5 into two families: writing down procedures and rules (skills, hooks, permissions), and splitting up and delegating work (subagents, agent teams). Each was tested in a small practice repository, with results checked on screen.

Within the delegation family, subagents and agent teams serve different needs. Subagents work in their own context and return only a summary; the article's measurement showed that handing over a 22-file research task cut what stayed in the conversation. Agent teams go further: the first conversation becomes a coordinator, launches roles, and lets those roles talk directly. The splitBill example showed the test agent writing tests from specs alone, never reading the implementation, and passing. But the article notes teams carry coordination overhead and use significantly more tokens, so the official guidance is to check whether a lighter method suffices first.

The stakes here likely hinge on how much coordination between roles a task actually needs. If direct back-and-forth between builder and tester is not required, subagents appear to deliver the benefit at lower cost. For teams already automating coding and testing with Claude Code, the practical question may be where the token budget pays off — and whether an experimental feature becomes a standard one.

FAQ
How do you turn on agent teams in Claude Code?
You set the environment variable CLAUDE_CODE_EXPERIMENTAL_AGENT_TEAMS to "1" in your .claude/settings.json file. The feature is experimental and off by default.
Do agent teams cost more than using Claude Code normally?
Yes. The article says teams use significantly more tokens than a single conversation because of coordination overhead, and official guidance recommends trying lighter methods like subagents first.
What did the test agent do in the agent teams example?
The test agent wrote tests in test/split.test.js for the splitBill function without ever reading the implementation, relying only on specs the builder agent provided. The tests passed.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleSwarmchasers: 400+ track AI agents' online traces