
What happened
Baseten, an AI infrastructure platform, is now integrated into Hugging Face Hub as a supported Inference Provider. Users can access Baseten-hosted models directly through the Hub's website, Python and JavaScript SDKs, and integrated agent tools. Initial support covers conversational and text-generation tasks, with models including DeepSeek V4 Flash, Kimi K3, and GLM-5.2.
Why it matters
Developers can now run a broader range of open-weight LLMs and other AI models through Hugging Face without setting up separate infrastructure. Users can choose between routing requests through Hugging Face (billed to their HF account with no added markup) or using their own Baseten API key. Hugging Face PRO subscribers receive $2 in monthly Inference credits usable across providers.
What to watch
Support for additional task types beyond text generation will roll out soon. The full list of Baseten-supported models is available at https://huggingface.co/baseten, and users can customize provider preferences in their account settings.
Summaries like this, in your inbox every morning.
Hugging Face Hub has expanded its ecosystem of Inference Providers to include Baseten, an AI infrastructure platform offering serverless AI and training services. This integration allows developers to access a wider range of models—particularly open-weight LLMs—without the friction of setting up separate accounts and infrastructure. The body describes two pathways for using Baseten through Hugging Face: direct billing (using a personal Baseten API key) or routed billing (through the Hugging Face account with no added markup). This dual approach removes barriers for both casual users and those with existing provider relationships.
The integration spans multiple entry points: the website UI (where providers appear on model pages ranked by user preference), client SDKs in Python and JavaScript, and popular agent harnesses including Pi, OpenCode, Hermes Agents, and OpenClaw. This breadth of integration points suggests that Hugging Face is positioning itself as a unified interface where developers can compare and switch between multiple inference providers without rewriting application code. The mention that Baseten will add support for additional task types (beyond the initial conversational and text-generation tasks) indicates this partnership is intended to deepen over time.
For example, today's edition would include:
AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
NVIDIA announced the NVIDIA Open Agent Safety Platform, an open software platform and reference design to secu…

NVIDIA announced its NVIDIA Open Agent Safety Platform, combining OpenShell open-source software for secure ag…

The Financial Times reports that Corporate America is embracing cheaper open AI models

Nvidia combined OpenShell, its March open-source sandbox software, with Sentry, a hardware watchdog for its Bl…

AWS open-sourced Strands Harness, built on the Strands Harness SDK it released in August 2026

Nvidia announced Monday its Open Agent Safety Platform, built on the OpenShell open-source software and the Ve…
