AIToday
AI Business & IndustryTechCrunch AIPublished: Jul 8, 2026, 19:00 JST2 min read

French startup ZML releases free LLM inference software for multi-chip AI deployment

French startup ZML releases free LLM inference software for multi-chip AI deployment

Key takeaway

  • ZML, a French AI startup, has released free inference software that lets organizations run AI language models across multiple chip types—Nvidia, AMD, Google, Apple, Intel—instead of being locked into a single vendor.

  • The software is designed to lower AI costs and energy consumption by giving enterprises flexibility in chip choice, and could help smaller European chipmakers compete with Nvidia's market dominance.

3 Key Points

  1. What happened

    ZML, a Paris-based AI startup backed by Turing Award winner Yann LeCun, has launched ZML/LLMD, a free inference server that allows open-source large language models to run on multiple chip types—including Nvidia, AMD, Google TPU, Apple Metal, and Intel Arc. The startup raised $20 million(約32億円) from venture firms including 20VC, >commit, AALVC, Drysdale Ventures, Kima Ventures, Kindred Capital, LocalGlobe, and Puzzle Ventures.

  2. Why it matters

    The software aims to break vendor lock-in and reduce AI costs by letting enterprises and cloud providers mix chips—some of which may be cheaper or more energy-efficient. This matters for businesses grappling with rising inference costs, since optimizing how AI systems process prompts has become more important than model training. The move could help smaller European chipmakers gain ground against Nvidia's dominance.

  3. What to watch

    ZML/LLMD is launching as a free product with no announced timeline for paid pricing; the company plans to learn about usage patterns before monetizing. The startup has a lean team of 20 people and founder Steeve Morin has flagged more releases to come.

Ask the AI about this article →

Context & Analysis

Inference—the step where an AI system produces an answer to a prompt—has become the dominant cost and performance bottleneck in AI deployment, even surpassing model training in importance. ZML is entering a crowded market that includes competitors like Baseten (valued at $13 billion(約2.1兆円)), Inferact (from the creators of open-source project vLLM), and RadixArk (the commercial entity behind SGLang). However, Morin's ambition extends beyond competing with vLLM and SGLang on narrow tasks; he frames ZML's mission as co-designing silicon with chipmakers and giving enterprises genuine optionality in their hardware mix.

The startup's ability to move quickly despite a small team of 20 people reflects both its founder's track record and healthy venture backing—a validation that angel investors and prominent founders see value in addressing vendor lock-in. Notably, the cap table includes Docker and Dagger founder Solomon Hykes, executives from Hugging Face (Clément Delangue and Julien Chaumond), and Yann LeCun, suggesting confidence that Europe can develop competitive AI infrastructure from home. By launching as a free product rather than immediately charging, ZML is prioritizing market adoption and understanding usage before monetizing—a deliberate strategy Morin articulated to avoid alienating early users.

FAQ

How much does ZML/LLMD cost?
ZML/LLMD is launching as a free product. Founder Steeve Morin said the company plans to learn about usage patterns before deciding when and how to charge for it.
Which chips does ZML/LLMD support?
The software works with Nvidia, AMD, Google TPU, Apple Metal, and Intel Arc chips, as well as open-source large language models across these different hardware types.
Who founded ZML and what is their background?
Steeve Morin founded ZML; he previously served as VP of engineering at Zenly, which Snapchat acquired for nine figures in 2017.

Get the latest AI Business & Industry news every morning

For example, today's edition would include:

  • Phonely launches Alma, voice AI trained on 10M callsSiliconANGLE AI · 49m ago
  • Aranya raises $11M to turn bare-metal servers into AI clusters in 48 hoursSiliconANGLE AI · 49m ago
  • CBTS launches Forge Agents for custom AI agentsSiliconANGLE AI · 49m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleFormer DeepMind Policy Chief Warns AI Arms Race Framing Risks Global Cooperation