AIToday
Large Language ModelsOpen-Source AIAI Business & IndustryArs Technica AIPublished: Sep 15, 2026, 22:00 JST

Mozilla: open models trail frontier AI by just 4.4 months

Mozilla: open models trail frontier AI by just 4.4 months

3 Key Points

  1. What happened

    Mozilla's State of Open Source AI report, published September 15, says the gap between top closed US frontier models and the best open-weights models has shrunk to 4.4 months.

  2. Why it matters

    Moonshot AI's Kimi K3 lands just three points behind Anthropic's Fable 5 on the Artificial Analysis Intelligence Index at 30 percent of the cost, and DoorDash already runs Kimi for routine work.

  3. What to watch

    Krikorian frames the decision as workload-specific, not organization-specific — paying for closed buys about a four-month head start at roughly five times the per-task cost, and only for tasks taking 8 to 12 hours.

WHO IT HITSEnterprise IT and platform teams choosing models for production workloads, plus finance and procurement groups weighing per-task model costs, face a genuine default-versus-premium decision rather than a blanket vendor choice. Developers without in-house staff to run open-weights models may still lean closed.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

Mozilla's report, previewed by Ars before its September 15 publication, lands as the second edition of its State of Open Source AI series, following an inaugural report on July 14. The gap it describes is measured several ways, including METR's "time horizon" — how long a task, in expert human hours, a model can complete with a reliable 50 percent success rate. Krikorian puts the current ratio at 1.7×: if open handles a seven-hour job, closed handles a 12-hour one, and in four months open catches up to 12 while closed reaches around 20. Vals AI's neutral-harness testing on Terminal-Bench 2.1 adds a cost dimension, with Z.ai's GLM 5.2 scoring within a point of Anthropic's Claude Opus 4.7 and 4.8 at about five times less per completed task.

The report pairs that capability convergence with a lopsided business picture. On OpenRouter, eight of the top 10 models by token volume in August 2026 offer open weights, yet a Linux Foundation paper by Frank Nagle and Daniel Yue found open models earning just 4 percent of overall revenue versus 96 percent for closed, based on data from May through September of 2025. Krikorian expects that revenue split to have shifted, and points to the geographic concentration behind today's open ecosystem: most open models the world runs on are Chinese, which he compares to the Android playbook of giving software away while owning the surrounding ecosystem.

Krikorian's proposed answer is a coalition — public compute programs, neutral foundations, companies that benefit from commodity models, and philanthropy — modeled loosely on how open source infrastructure like Linux was funded. The stakes hinge on whether such an "alternative coalition" of mission-driven institutions, rather than frontier labs, can materialize before China's position in open models hardens further.

FAQ
How far behind are open models compared with closed frontier models?
According to Mozilla's report, the performance gap has closed to just 4.4 months. Moonshot AI's Kimi K3 scores only three points behind Anthropic's Fable 5 on the Artificial Analysis Intelligence Index.
When are closed frontier models still worth paying for?
Mozilla CTO Raffi Krikorian says closed models earn their premium for expert professional work, high-intensity retrieval, long context, and tasks taking 8 to 12 hours. That head start costs roughly five times more per completed task.
Who is already using open models for real work?
DoorDash has been using Kimi for routine work while reserving Fable for harder tasks. Mozilla's report also notes eight of the top 10 models by token volume on OpenRouter in August 2026 provide open weights.
Ars Technica AIRead Original Article

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Ode CEO Chris Taylor: frontier AI slowdown won't slow big firmsSemafor Tech · 3h ago
  • Meta One bundles start at $7.99/month with AI Muse perksThe Verge AI · 3h ago
  • AIUC raises $40 million to audit rogue AI agentsTechCrunch AI · 3h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleWachter: AI data centers need 2.7× productivity by 2030