Microsoft puts 3M Hugging Face models on GPU tap, no Dockerfile required
Hugging Face Blog · Jul 7, 2026 · 2 min read
Microsoft is cracking open the operational bottleneck that's kept millions of Hugging Face models out of production. Th...
8 articles
Explore our coverage of model deployment — 8 curated articles, summaries, and related resources from the Inblix archive.
Hugging Face Blog · Jul 7, 2026 · 2 min read
Microsoft is cracking open the operational bottleneck that's kept millions of Hugging Face models out of production. Th...
Hugging Face Blog · Nov 13, 2025 · 2 min read
The open model library Hugging Face and Google Cloud are deepening their tie-up with a concrete engineering fix for a m...
Hugging Face Blog · Jun 12, 2025 · 2 min read
Hugging Face just plugged a major gap in its inference ecosystem by onboarding Featherless AI as a supported provider....
Hugging Face Blog · May 19, 2025 · 2 min read
Two years ago, Microsoft and Hugging Face began a modest experiment: make open models easier to find and deploy on Azur...
Hugging Face Blog · Jan 28, 2025 · 2 min read
Hugging Face is finally opening up its serverless Inference API to third-party providers, a move that acknowledges a si...
Hugging Face Blog · Jan 22, 2025 · 2 min read
Hugging Face just gave its users a faster on-ramp to production AI. A new partnership with FriendliAI—ranked by Artific...
Hugging Face Blog · Dec 9, 2024 · 2 min read
Amazon just tore down a big wall between its walled garden and the open-source AI world. As of today, 83 popular open m...
Hugging Face Blog · Aug 19, 2024 · 2 min read
If you want to run Meta's monster 405-billion-parameter Llama 3.1 model without it melting your hardware, Google Cloud'...