
What happened
In a repository running 7 subagents, qa-reviewer — limited to read-only tools — rejected a draft because a first-person anecdote lacked a verification record; after evidence was logged, a re-review passed.
Why it matters
Separating the writer from the checker appears to work, since a reviewer that cannot edit files caught an unsupported claim the author would likely have passed themselves.
What to watch
builder and content-writer were merged for two tightly coupled articles, and growth-marketer and revenue-analyst have never run, so the count of 7 may be an overestimate.
WHO IT HITSEngineers and solo developers configuring AI coding agents (Claude Code subagents) now have a concrete split criterion — different output location or different evaluation standard — plus evidence that a read-only checker role catches unverified claims. Teams copying agent setups between projects should note that unused roles are cheap to keep, since each definition is a single file under .claude/agents/.
Summaries like this, in your inbox every morning.
This is a first-day field report rather than a proven playbook. The seven subagents were split along three design criteria: whether the output destination differs, whether the required expertise and judgment criteria differ, and whether the role is worth reusing independently. The author is explicit that only the first of these is mechanically checkable — you can look at the directories — while whether two jobs really use different evaluation standards often only becomes clear after running both in one agent. The qa-reviewer separation is the case where that second criterion was confirmed the hard way: a defect appeared, and only then was it clear the evaluation axes had genuinely differed.
Within the same repository, the opposite happened at the same time. builder and content-writer were defined as separate roles with different output folders (samples/ and articles/) and different expertise, yet for two articles whose sample code and prose were one-to-one coupled, running them as one agent was faster because no other judgment intervened between the steps. Meanwhile growth-marketer and revenue-analyst have never been invoked; the ops/kpi.md update that revenue-analyst was meant to own amounted to the main session adding a single line after one article was published.
What ties these together is that agent definitions live as Markdown files under .claude/agents/, so merging or splitting costs one or two file edits. That low correction cost is arguably the reason the author can leave the question open — and the reason a follow-up review after roughly a month of operation is planned to test whether 7 was the right number.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
Microsoft launched MAI-Transcribe-2-Streaming, its first streaming transcription model, priced at 54 cents per…
Google announced Gemini 4 Argon on September 30, saying DeepMind's own evaluation beat GPT-6 Astra, Claude Fab…

Microsoft AI said it released MAI-Transcribe-2-Streaming, which returns provisional results in just over 100 m…

Claude Code now lets users change its behavior, customize the UI, and swap in their own features using a few l…

Anthropic said it made the web service claude.ai and its desktop app about 3 times faster in 2 weeks, and that…

On October 1, OpenAI updated ChatGPT's release notes with shopping features — a 'try on' button on product car…
