
A new Princeton-led study found that AI agents cannot yet conduct open-ended research, even though they excel at engineering tasks.
Researchers tested Claude Opus 4.8 on unpublished research questions from top-tier machine-learning conference papers, but both resulting papers were rejected for lack of originality and poor judgment.
The finding challenges industry hype about rapid recursive self-improvement—where AI systems improve themselves—and suggests that creative, exploratory thinking remains beyond current AI capabilities.
What happened
Princeton researchers tested whether AI agents could conduct open-ended research—the kind without clear-cut answers that requires judgment and creativity. They had Claude Opus 4.8 tackle unpublished research questions from NeurIPS 2026 papers over six days with $3,000 in API credits. Both papers were rejected by the original authors; the agents handled engineering tasks well but produced work that lacked novelty and failed to explore ideas properly.
Why it matters
The AI industry has promised that models will soon improve themselves with minimal human oversight, but this study suggests that milestone is further away than many forecasts claim. AI systems can write code and optimize chips, but they struggle with the open-ended thinking—rethinking approaches, exploring multiple ideas, incorporating feedback—that actual research progress requires. Anthropic cofounder Jack Clark called the lack of creativity a "bearish signal on short recursive self-improvement timelines."
What to watch
The research team is now testing the same experiment with Mythos, Anthropic's most advanced model launched in April, though it is now available only to approved organizations under Trump administration safety restrictions. The core question remains unresolved: whether recursive self-improvement requires creative breakthroughs or can rely solely on narrower, measurable improvements to speed and benchmarks.
Ask the AI about this article →
The promise of recursive self-improvement—where AI systems accelerate their own development—has become the AI industry's boldest near-term milestone. Both OpenAI and Anthropic have publicly charted progress toward this goal. In July, OpenAI highlighted that GPT-5.6 Sol helped post-train a smaller model, saving weeks of work. Anthropic published a blog post in June titled "When AI Builds Itself," signaling confidence in the trajectory. However, the Princeton study exposes a critical gap between what AI agents can do narrowly and what they can do creatively.
The researchers designed "shadow evaluation," a new testing method, to measure open-ended research capabilities—the kind of work that cannot be reduced to yes-or-no answers. Their finding that Claude Opus 4.8 could manage all the engineering legwork (literature review, experiment running, result compilation) but produced papers rejected for lack of originality points to a deeper limitation. Sayash Kapoor, one of the study's leads, attributed this to how AI models are trained: they excel at tasks that can be checked automatically through reinforcement learning, but open-ended work requires training environments that do not yet exist. Anthropic cofounder Jack Clark later echoed this observation, noting the company found similar constraints when attempting to automate AI safety research—a domain also demanding intuitive judgment over rote execution.
The study has limits. It evaluated only two papers, and knowing their graders were human researchers evaluating AI work may have influenced their judgments. The researchers themselves had discretion in study design, introducing potential bias. Yet the findings align with what AI companies report internally and raise a structural question about the path to self-improvement: can AI progress through grinding improvements on narrow, measurable tasks alone, or does recursive self-improvement require the creative leaps—like the invention of transformers—that Kapoor argues have driven the field's biggest breakthroughs? That, Kapoor concludes, is "frankly the trillion-dollar question right now."
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
Ask AI anything about this article. Q&As are published on this page for other readers too.
HP Korea has formed a partnership with Upstage, a large language model (LLM) startup, to advance its localized…

At the "AI on Chips: Semiconductor Industry Trends Forum" hosted by DIGITIMES, industry experts highlighted th…

Anthropic launched Claude Academy on August 20, a free learning site that explains AI fundamentals and how to…

OpenAI rolled out support on Thursday for controlling Apple's iMessage service via ChatGPT on Mac, enabling th…

As AI technology matures, the bottleneck in the industry is moving beyond semiconductor constraints like GPUs…

OpenAI has launched an Apple Messages plug-in for ChatGPT that lets users connect their Messages inbox to the…
