AIToday
Large Language ModelsOpen-Source AIHugging Face BlogPublished: Oct 8, 2026, 04:00 JST

Liquid AI's d1-3B tops Decision Index 0.2.1 at 48.57

Liquid AI's d1-3B tops Decision Index 0.2.1 at 48.57

3 Key Points

  1. What happened

    Liquid AI released d1-3B, which scores 48.57 on the Decision Index 0.2.1 — ahead of every 4B and 9B model and of Decider 35B-A3B (47.11).

  2. Why it matters

    A small open model can now outrank much larger rivals on decision tasks, meaning teams can consider running it on their own hardware rather than paying for a bigger model.

  3. What to watch

    The edge speed claims hinge on the specific NVIDIA hardware measured, and the companion d1-omni-600M is an early research release with no speed numbers or vision/audio benchmark scores yet.

WHO IT HITSDevelopers and device makers building on-device decision features — such as customer-ticket routing or image checks — can now evaluate a small open model that runs on NVIDIA's edge hardware; whether it fits a specific product depends on their latency and footprint needs.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Liquid AI built these two models on its Liquid Foundation Models, but with a twist: unlike generative models that produce tokens, decision models answer in a single forward pass. That design choice is what makes the edge-speed numbers possible. d1-3B is trained from LFM2.5-VL-3B, a decoder-only vision-language model, while the experimental d1-omni-600M is trained from LFM2.5-Encoder-350M, a bidirectional encoder that adds vision and audio encoders to handle all three modalities. The two backbones point at different use cases: d1-3B for text-and-image decisions where quality matters, and d1-omni-600M where footprint is the constraint.

The benchmark results show d1-3B leading on mean score across seven public datasets covering reading comprehension, toxicity detection, intent classification, medical QA, and cross-lingual understanding. The company did not report vision or audio benchmarks, noting that the Decision Index v0.3 includes only a private vision split and that audio decision benchmarks are currently an open problem. That gap matters for anyone hoping to compare the multimodal claims directly against rivals.

The speed validation, done with NVIDIA across GeForce RTX 4090, Jetson AGX Thor, Jetson AGX Orin 64 GB, and Jetson Orin Nano, shows d1-3B answering a single question in under 50 ms on every measured device. Whether that translates into real-world product wins will likely depend on how the models perform on each team's specific decision tasks, and on when d1-omni-600M emerges from its early research release.

FAQ
How fast is d1-3B on edge devices?
Liquid AI says d1-3B answers a question in 16 ms on an NVIDIA Jetson AGX Thor and 50 ms on a Jetson Orin Nano. Three questions take only 1.3x the time of one on these devices.
What inputs can the two models handle?
d1-3B supports text and images, while d1-omni-600M supports text and images or text and audio. Both are open-weight and available on Hugging Face.
What are the benchmark scores?
On seven public datasets, d1-3B has a mean score of 82.9 and d1-omni-600M scores 78.4, surpassing Decider 2B (77.1) with only a quarter of the parameters.
Hugging Face BlogRead Original Article

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleOpenAI to bring College Planner to ChatGPT for Teens