
Broadcom announced VMware AI Factory at VMware Explore 2026.
It automates AI infrastructure deployment and management.
This reduces time to first AI model from weeks to hours.
What happened
Broadcom announced VMware AI Factory, a software-defined foundation for VMware Private AI Cloud, at VMware Explore 2026. It automates AI infrastructure deployment and lifecycle management.
Why it matters
This cuts the time from bare metal server deployment to serving the first AI model from weeks to hours. It also pools GPU resources so teams can run multiple models on shared hardware, lowering costs and improving tokenomics.
What to watch
Broadcom is partnering with MetalSoft to reduce bare-metal provisioning from weeks to minutes. It also collaborates with AMD to support AMD Instinct GPUs and the open ROCm software ecosystem.
Ask the AI about this article →
The announcement addresses a key pain point for enterprises: the slow, complex journey from hardware to running AI models. By automating infrastructure provisioning and lifecycle management, Broadcom aims to make private AI more practical and cost-effective, directly tackling the 'metal to model' challenge highlighted by Paul Turner.
The pooling of GPU resources and unified model governance through services like Multi-tenant Model Sharing and AI Gateway suggests a focus on operational efficiency and cost control. The partnerships with AMD, Cisco, Lenovo, Supermicro, and MetalSoft indicate an effort to provide customers with flexibility and choice, avoiding lock-in to a single vendor stack.
This move positions Broadcom to compete in the private AI infrastructure space by emphasizing speed, automation, and predictability. The emphasis on data sovereignty and running AI where data lives is likely to appeal to enterprises with strict compliance or latency requirements.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
AI system scaling has pushed interconnect requirements inside data centers from chips and boards up to racks…

Chinese large-model developer Z.ai says it can now support large-scale inference using roughly 100,000 domesti…

Analyst Ming-Chi Kuo says Nvidia has revived the Rubin CPX AI accelerator with a substantially redesigned arch…

Palantir Technologies stock has posted multi-year gains, including an 11x return over 3 years

Apple has escalated its legal battle against OpenAI, claiming in a new court filing that OpenAI is actively de…

Samsung Electronics has locked up as much as 70% of its memory production capacity under long-term supply agre…
