
Templar's Crucible, a distributed pre-training platform with data-parallel replicas and pipeline parallelism, simulated stage skipping on a 178M model with eight replicas and four stages per replica. At a 1% per-replica failure probability per global step, validation loss stayed close to the no-failure baseline even when each outage removed a stage for six global steps.
Summaries like this, in your inbox every morning.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, LINE, or Slack.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Anthropic released Claude Opus 5.5 today and cut its price 20%, with input at $4 per million tokens and output…
Anthropic and OpenAI, which spent early September warning that model capabilities are outrunning the safeguard…

OpenAI released GPT-6 Sol and GPT-6 Luna on September 22, trained the same way as its top-tier GPT-6 Astra, an…

In September 2026, Amazon Web Services and Salesforce announced that Salesforce CRM data will be usable from A…

Anthropic announced Claude Opus 5.5 on September 22, the first model in its new Claude 5.5 series, and began o…

John Platt and his Google team built Empirical Research Assistance (ERA)
