AIToday
AI Business & IndustrySiliconANGLE AIPublished: Oct 7, 2026, 13:01 JST

Caterpillar and CoreWeave cut physical-AI learning loop to hours

Caterpillar and CoreWeave cut physical-AI learning loop to hours

3 Key Points

  1. What happened

    Caterpillar and CoreWeave shortened the data-labeling and feedback loop for training autonomous construction machines from months to maybe weeks, and now hours within a given workday.

  2. Why it matters

    That speed could let Caterpillar's roughly 18 petabytes of federated machine data be turned into usable simulation and training input the same day, easing the shortage of skilled operators.

  3. What to watch

    Hootman still calls the 18 petabytes a drop in the bucket, so the test is whether hours-long loops hold as data grows; CoreWeave named Caterpillar a customer only in its second-quarter results.

WHO IT HITSThis lands on construction and mining firms looking to run autonomous equipment, whose productivity gains hinge on how fast field data becomes training-ready. Engineers building physical-AI systems at those firms are the ones who would adopt or compete with this embedded-service model.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The conversation at the Fully Connected event framed construction as a harder problem than mining, where Caterpillar has years of experience with autonomous equipment. Brandon Hootman argued that a mine site changes infrequently, while construction is its polar opposite, requiring systems that match a structured setup to an unstructured environment. Richard Ahlfeld added that physical AI is an entirely different beast from the AI clouds built for foundation-model training and agentic inference, requiring a lot of storage and a different infrastructure.

Caterpillar began working with CoreWeave this year, drawn by both graphics processing unit capacity and applied expertise, and CoreWeave later named Caterpillar among its enterprise customers in its second-quarter results announcement. Working with Nvidia, the partners use AI models to annotate and label incoming field data. Training an autonomous excavator means ingesting telemetry and vision data, simulating a digging scenario a million times and adding reinforcement learning. Hootman noted that a single machine can produce terabytes of data within a given day, encompassing Light Detection and Ranging data, camera data, multi-second control data and performance data.

Whether the hours-long loop holds may depend on how those terabytes accumulate, since Hootman described the current 18 petabytes as only a drop in the bucket. For construction and mining firms facing declining productivity and fewer skilled operators, the value of this approach hinges on whether that accelerated loop survives that scale, and on whether CoreWeave's embedded-engineer model proves repeatable beyond this partnership.

FAQ
How much machine data does Caterpillar have?
Caterpillar's digital ecosystem holds about 18 petabytes of federated data from machines, dealers and customers. Brandon Hootman says that is still a drop in the bucket.
What service did CoreWeave launch for physical AI?
CoreWeave recently launched a Physical AI Field Engineering service that embeds its engineers with customers' domain experts.
SiliconANGLE AIRead Original Article

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleMistral Large 4 preview live; open weights due late October