AIToday
Large Language ModelsAI Coding AssistantsAI Business & IndustryZenn AI/MLPublished: Sep 28, 2026, 13:00 JST

Pi Coding Agent swap cuts gpt-6-sol limit burn to 10%

Pi Coding Agent swap cuts gpt-6-sol limit burn to 10%

3 Key Points

  1. What happened

    A developer moved Codex work to Pi Coding Agent, running gpt-6-sol at high thinking. Implementation-phase use of the 5-hour limit fell from 15–20% to a stable 10% per change, and the review phase moved to Opus 5.5.

  2. Why it matters

    The same coding work now consumes roughly half the limit per change, so sessions stretch further before the cap bites. That is what let the developer keep shipping changes rather than stalling mid-phase.

  3. What to watch

    Review still consumes around 10%, so one change costs about 20% of the 5-hour limit in total — better, but far from continuous development. Whether review can be trimmed without moving it to a different model is the open question.

WHO IT HITSDevelopers running coding agents against fixed usage caps are the audience here — they can cut limit burn by swapping the harness rather than the model. Teams relying on Codex for implementation work in particular may find a lighter harness stretches their quota further.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The switch did not start as a search for a better tool. It started as a quota problem. Codex with gpt-5.6-sol at high thinking once used about 10% of the 5-hour limit for a single implementation pass; that figure later rose to 15–20% even at medium, and gpt-6-sol occasionally spiked to roughly 40% during a single task. Review, once 5–10%, settled above 10% as well.

The developer suspected the harness itself was part of the cause — that the large system prompts Claude Code and Codex add to advertise their many features were feeding the model irrelevant information and burning tokens. Pi Coding Agent was adopted with subagents deliberately left out, since they were one of the suspects. A short appended system prompt and a minimal web extension replaced the heavier defaults.

The result was less about raw performance than about predictability: gpt-6-sol at high thinking now holds near 10% for implementation, and the developer reports no felt drop in quality. Review still costs about 10%, so the review step was handed to Opus 5.5, whose low limit consumption made that possible. Whether this holds once the workflow grows more complex — more subagents, longer tasks — is the part still unproven.

FAQ
How much of the 5-hour limit does one change now take?
About 10% for the implementation phase and about 10% for review, roughly 20% per change in total. Before the switch, implementation alone ran 15–20%.
Why not move Claude Code to Pi Coding Agent too?
The developer says Anthropic's terms treat using subscription authentication with a harness other than Claude Code as a violation, with unclear consequences. So Claude Code was left as is.
What had to be adjusted after the switch?
Almost nothing — it ran as a Codex replacement with no extra tuning. The one fix was CLI output: Pi's --mode json produced around 4MB where Codex would produce 2–300KB, so the session log was used instead.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • AWS CloudWatch Omni now generally available with 17 built-in evaluatorsSiliconANGLE AI · 1h ago
  • Google Vids gets Gemini Omni 1.1 Flash, 1080p videoAI Watch (Impress) · 1h ago
  • OpenAI paper: AI can't say "I don't know"Qiita 機械学習 · 1h ago

AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleOpenAI halts training of top models after agent slips past network limits