
OpenAI's Astra model worries AI safety experts.
It appears to reason without showing its steps.
The debate follows a hack at Hugging Face.
What happened
OpenAI's new AI model Astra is delighting fans by completing tasks with very little human intervention, but AI safety experts are concerned about the lack of visibility into its reasoning.
Why it matters
Researcher Ryan Greenblatt called Astra's ability to solve hard competition math problems 'extremely concerning,' especially after the Hugging Face hack exposed limits in understanding AI models' 'chain of thought.'
What to watch
OpenAI's chief scientist Jakub Pachocki has pushed back against reports of purposely limited visibility, saying he wants to prevent 'a race into unmonitorability.'
Ask the AI about this article →
The article highlights a tension between AI capabilities and oversight. Astra's efficiency—solving complex problems with minimal visible reasoning—appeals to users but alarms researchers who need to understand AI decision-making. The recent Hugging Face hack adds urgency, showing that even experts cannot fully grasp AI's internal logic. OpenAI's chief scientist has responded to reports, but his comments aim to temper fears rather than resolve them. As AI models become more autonomous, the industry faces a challenge: how to ensure safety without sacrificing performance.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
LINE Yahoo is expanding ad delivery using its AI agent 'Agent i'

Furukawa Electric is a top supplier of external laser sources (ELS) for AI data center CPO switches, with high…

Tokyu Construction announced on August 31, 2026, that it will use NTT ConoSurf's voice AI and generative AI to…

Mitsubishi Heavy Industries and NEC announced on the 2nd that they will strengthen cooperation in the defense…

Sony Group and 35 other companies have filed a lawsuit against Anthropic in the U.S., according to the article…

Meta formally ended an incentive program where employee performance evaluations depended on how much they used…
