AIToday
Large Language ModelsAI Safety & AlignmentTHE DECODERPublished: Aug 30, 2026, 22:01 JST2 min read

AI agents have no sense of time, study finds

AI agents have no sense of time, study finds

Key takeaway

  • AI coding assistants cannot judge how long tasks take.

  • They overestimate time needed, especially for short jobs.

  • They also overrate their own work quality.

3 Key Points

  1. What happened

    A study by two independent AI researchers, done as part of the MATS research program, tested Anthropic's Claude Code and OpenAI's Codex on their sense of time. Before tasks, the agents had to estimate how long they'd need; afterward, they reported how much time had passed. The test material came from 200 tasks in ProgramBench plus 18 custom benchmarks.

  2. Why it matters

    The agents consistently overestimated how much time they'd need. On ProgramBench, both models mostly guessed around 90 minutes, no matter the difficulty. In the second round, Claude was off by three times on average, Codex by six to ten times. The estimates were worst for short tasks, and only in the multi-hour range did some predictions come close to reality.

  3. What to watch

    The agents also overrated their own work: older models Opus 4.8 and GPT-5.5 scored themselves about 20 points higher on average, even on failed tasks. When the agents got access to a tool that reports elapsed time, they got it right almost every time.

Ask the AI about this article →

Context & Analysis

The study highlights a fundamental limitation of AI agents: they lack an internal clock. This matters because long-running tasks require agents to follow instructions like 'iterate for two hours,' which becomes impossible if they cannot gauge the passage of time. The findings also show that runtime depends heavily on the surrounding software (the harness), with the same model taking 2.5 times more steps in Claude Code than in Codex on average. The researchers plan to test whether agents can stick to a set work duration, suggesting a potential direction for improving agent reliability.

FAQ

Why can't AI agents sense time?
The study found that coding assistants like Claude Code and Codex lack a reliable sense of time, leading them to overestimate task duration and misjudge how long they've been working.
What happens when agents get a time-reporting tool?
When the agents had access to a tool that reported elapsed time, they got the timing right almost every time, according to the study.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • JetBrains launches free local coding agent Junie LocalPublickey · 17m ago
  • 34% of Americans Use AI Chatbots for Health: PewHacker News · 17m ago
  • Caterpillar turns mining automation expertise to AI deploymentTechCrunch AI · 17m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleS. Korea's aging population may blunt AI boom wealth