AIToday
Open-Source AIHugging Face BlogPublished: Apr 30, 2026, 04:00 JST1 min read

DeepInfra is now a supported Inference Provider on the Hugging Face Hub, offering serverless AI inference with support for conversational and text-generation tasks.

DeepInfra is now a supported Inference Provider on the Hugging Face Hub, offering serverless AI inference with support for conversational and text-generation tasks.

3 Key Points

  1. DeepInfra joins Hugging Face's ecosystem of Inference Providers, enabling developers to access models like DeepSeek V4, Kimi-K2.6, and GLM-5.1 directly from model pages and through client SDKs (Python and JavaScript).

  2. Users can authenticate via Hugging Face token for routed requests (billed at standard provider rates with no markup) or use their own DeepInfra API key for direct requests (billed by DeepInfra). Hugging Face PRO users receive $2 worth of Inference credits monthly across providers.

  3. Initial launch supports conversational and text-generation tasks on open-weight LLMs; support for text-to-image, text-to-video, embeddings, and other tasks will roll out soon.

Ask the AI about this article →

Hugging Face BlogRead Original Article

Get the latest Open-Source AI news every morning

For example, today's edition would include:

  • Hugging Face launches 207 WebGPU kernels for fast in-browser AIHugging Face Blog · 48m ago
  • DataAgent launches with $10M to auto-fix Kubernetes faultsSiliconANGLE AI · 6h ago
  • Z.ai runs GLM on 100,000 Chinese AI chipsDIGITIMES Asia · 9h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleApple researchers introduce Sonata, a method that adaptively allocates thinking budgets to large language models, achieving 20% to 80% reduction in thinking tokens while maintaining accuracy.