AIToday
Large Language ModelsTop Companies' AI MovesAI Business & IndustryTop Companies AIPublished: Sep 17, 2026, 06:30 JST

Intel Xeon gains 2.4x in MLPerf v6.1 via software

Intel Xeon gains 2.4x in MLPerf v6.1 via software

3 Key Points

  1. What happened

    On the same Xeon 6980P silicon and socket count as MLPerf v6.0, Intel reported a 2.4x rise in Llama 3.1 8B Server throughput and 56% higher Offline throughput in v6.1, achieved through software alone.

  2. Why it matters

    Intel says optimizations can stretch the useful life of servers customers already run, letting them handle larger, more varied AI models without buying new hardware.

  3. What to watch

    The gains apply to specific submitted configurations, so real-world results will hinge on each customer's setup and software stack. Partner submissions rose from 29 to 39, with Oracle, Red Hat, Quanta Cloud Technology and Supermicro joining.

WHO IT HITSEnterprise infrastructure teams running Xeon-based servers can potentially defer hardware refreshes if they adopt the software updates. Buyers evaluating AI inference platforms may find Intel's growing partner validation useful, though results depend on their own configurations.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

MLPerf Inference is a widely watched set of benchmarks run by MLCommons that lets companies compare how fast different hardware handles AI tasks. In the v6.0 round, Intel's submissions focused on a narrower set of Xeon and Arc Pro configurations. For v6.1, Intel expanded its Xeon participation from two benchmarked SKUs to five, increasing CPU inference results from 24 to 35, and its Arc Pro B70 submissions now cover more models including Llama 2 70B and gpt-oss-120B.

The company also pointed to broader ecosystem involvement, with Oracle making its first Intel-based submission and Red Hat delivering its first Xeon CPU inference submission. Quanta Cloud Technology and Supermicro provided the first partner submissions using Intel Arc Pro B70 GPUs. Intel says its Xeon improvements are being upstreamed into widely used AI frameworks so customers can benefit on deployed infrastructure, and its Arc Pro B70 optimization work is intended to advance software for future Intel GPU products.

The stakes for Intel's argument rest on whether these software gains translate into real cost savings for businesses running AI inference. That likely depends on how closely a customer's setup matches the benchmarked configurations and whether they adopt the updated software. If the improvements hold in typical enterprise environments, they could help delay hardware replacement cycles, a meaningful consideration for IT teams managing tight budgets.

FAQ
How much did Intel's Xeon 6980P improve in MLPerf v6.1?
Intel reported a 2.4x increase in Llama 3.1 8B Server throughput and a 56% increase in Offline throughput compared to v6.0 on the same hardware.
Did Intel's partners contribute more results in v6.1?
Yes, third-party submissions on Intel platforms increased from 29 in v6.0 to 39 in v6.1, including first-time contributions from Oracle and Red Hat.
What new benchmark did Intel support in v6.1?
Intel co-developed and submitted results for the new end-to-end retrieval-augmented generation (E2E-RAG) benchmark, splitting workload between Xeon CPU and Arc Pro B70 GPUs.
Top Companies AIRead Original Article

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Mark Zuckerberg: labs ignoring "focus on alignment will fall behind"Top Companies AI · 1h ago
  • Apple's Smarter Siri Still Lags in AI Race, WSJTop Companies AI · 1h ago
  • Amazon Bedrock, Azure AI Foundry, Vertex AI diverge on priceTop Companies AI · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleSchwab brings Anthropic's Claude to 16,000 advisors