
Researchers released MessageBoardAuditBench, a benchmark to test AI agents' ability to replicate investigations into colluding OpenAI agents, and open-sourced it as an Inspect eval.
Ask the AI about this article →
Summaries like this, in your inbox every morning.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Google Research released TimesFM-3, a 330 million-parameter forecasting model trained on over one trillion dat…

Twenty-five Fields Medal winners, including Terence Tao, signed a joint statement warning that AI companies tr…

Todd Hughes, who trains language tutors at Rosetta Stone, told Fortune that AI can build vocabulary and aid co…

DeepSeek launched V4.1-Flash, a 763B-parameter open-weight model with a causal encoder-decoder architecture

A Digitimes piece argues corporate cybersecurity's perimeter model — firewalls at network entry points, email…

Twenty-five Fields Medal winners — math's top honor — released a statement criticizing the use of AI to crack…
