AIToday
Large Language ModelsAI Coding AssistantsITmedia AI+Published: Sep 18, 2026, 10:01 JST

Anthropic says 'just-in-case' prompts waste Claude tokens

Anthropic says 'just-in-case' prompts waste Claude tokens

3 Key Points

  1. What happened

    Anthropic published a blog post on September 8, 2026 (US time), outlining six common prompting anti-patterns that waste Claude tokens, and tested its new /claude-api prompt-audit command, which reduced costs by an average of 14.6% and improved accuracy by an average of 5.3%.

  2. Why it matters

    The company says the cost-versus-performance tradeoff is not inevitable; targeting just three areas—advanced tool use, prompt caching, and effort adjustment—can deliver both savings and maintained or improved performance, and it has packaged this guidance into a skill available in Claude Code.

  3. What to watch

    The recommendations hinge on prompt caching working as intended, so the specific configuration choices Anthropic describes—such as keeping the tools and system prompt unchanged, avoiding mid-conversation changes to the tools or system prompt, and placing static content at the beginning of the prompt—are likely to determine whether readers see similar gains. The article also notes that prompt cache expires, so cache efficiency may require continuous management.

WHO IT HITSEngineering and product teams building applications on Claude Platform—especially those already using Claude Code and its API—can apply the anti-pattern audit and the three focus areas to reduce token costs without sacrificing accuracy. Teams that maintain prompt caches will need to watch for configuration changes that invalidate the cache, which the body says can raise costs.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

Anthropic's blog post and the accompanying commands—/claude-api prompt-audit, /claude-api hillclimb, and /claude-api cost-optimize—reflect a push to make Claude more efficient at a time when model providers face pressure on both price and quality. The company's tests show that applying its prompt-audit command during a shift from Opus 4.8 to Opus 5 delivered both lower costs and higher accuracy, and that hillclimb explored model, effort level, and prompt updates to reach 98.9% accuracy on a training set at about 2.6 cents per thousand tokens. The guidance also acknowledges that prompt caching, while valuable, depends on exact byte-level prefix matches and a limited cache lifetime. For teams building on Claude Platform, the practical question is whether Anthropic's own test results transfer to their specific applications, since the body does not claim they will. The commands and skills are already available, so the main uncertainty is whether users can match the configuration conditions Anthropic describes for cache preservation and cost reduction.

FAQ
What is the /claude-api prompt-audit command?
It is a new Anthropic command for Claude Code that scans a project directory—prompts, skills, tools, Claude Code settings such as CLAUDE.md, and code that calls the Claude API—to find anti-patterns. Anthropic tested it with customers and reported an average 14.6% cost reduction and 5.3% accuracy improvement.
How does Anthropic recommend reducing Claude costs?
Anthropic suggests three main approaches: managing token usage through prompt caching and request structure, adjusting the effort level to match the task, and using the /claude-api cost-optimize command to generate reduction ideas based on usage reports and token consumption logs. It also recommends trying stronger models with lower effort settings.
What are the six prompting anti-patterns Anthropic identified?
They include unnecessary re-verification instructions (e.g., 'double-check the work'), emphasis markers (e.g., 'this is very important'), fixed procedure or chain-of-thought instructions (e.g., 'think step by step'), few-shot examples tuned for older model execution modes, redundant instructions, and settings already described in Claude's configuration such as thinking budgets.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • OpenAI launches Astra for Law, a GPT-6 setup for legal researchSiliconANGLE AI · 1h ago
  • Google opens CC to families of six as shared AI agentSiliconANGLE AI · 1h ago
  • Anthropic shares pacing metrics, cites 30,000 agentsSiliconANGLE AI · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articlePalantir AI deals meet $171.06 fair value debate