
DeepInfra joins Hugging Face's ecosystem of Inference Providers, enabling developers to access models like DeepSeek V4, Kimi-K2.6, and GLM-5.1 directly from model pages and through client SDKs (Python and JavaScript).
Users can authenticate via Hugging Face token for routed requests (billed at standard provider rates with no markup) or use their own DeepInfra API key for direct requests (billed by DeepInfra). Hugging Face PRO users receive $2 worth of Inference credits monthly across providers.
Initial launch supports conversational and text-generation tasks on open-weight LLMs; support for text-to-image, text-to-video, embeddings, and other tasks will roll out soon.
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Hugging Face released @huggingface/kernels, a library for running optimized WebGPU kernels from the Hugging Fa…

Israeli startup DataAgent Ltd
Chinese large-model developer Z.ai says it can now support large-scale inference using roughly 100,000 domesti…

Broadcom announced VMware AI Factory, a software-defined foundation for VMware Private AI Cloud, at VMware Exp…

OpenClaw launched version 2.0, its largest update yet, with a version number of 2026.8.1
David Heinemeier Hansson (DHH), creator of Ruby on Rails, has released Omarchy 4.0 (Omarchy Quattro), the late…
