AI Pulse by Inblix

Topic: serverless-inference

9 articles

Explore our coverage of serverless-inference — 9 curated articles, summaries, and related resources from the Inblix archive.

Baseten Joins Hugging Face Hub, Bringing DeepSeek V4 Flash to Serverless Inference — Inblix summary
Research

Baseten Joins Hugging Face Hub, Bringing DeepSeek V4 Flash to Serverless Inference

Hugging Face Blog · Aug 6, 2026 · 2 min read

Baseten is now a supported Inference Provider on the Hugging Face Hub, a move that plugs its serverless infrastructure...

DeepInfra joins Hugging Face Hub, slashing serverless AI costs for devs — Inblix summary
Research

DeepInfra joins Hugging Face Hub, slashing serverless AI costs for devs

Hugging Face Blog · Apr 29, 2026 · 2 min read

Hugging Face just got a serious cost-performance boost. DeepInfra, the serverless inference platform known for aggressi...

OVHcloud brings sub-200ms inference to Hugging Face, starting at €0.04 per million tokens — Inblix summary
Research

OVHcloud brings sub-200ms inference to Hugging Face, starting at €0.04 per million tokens

Hugging Face Blog · Nov 24, 2025 · 2 min read

Hugging Face is expanding its Inference Providers ecosystem with OVHcloud, bringing a distinctly European option to the...

Scaleway Brings Sub-200ms AI Inference to Hugging Face, Starting at €0.20/M Tokens — Inblix summary
Research

Scaleway Brings Sub-200ms AI Inference to Hugging Face, Starting at €0.20/M Tokens

Hugging Face Blog · Sep 19, 2025 · 2 min read

Hugging Face just got a serious European hardware upgrade. Scaleway is now a supported Inference Provider on the Hub, w...

Hugging Face adds Public AI as a free, nonprofit serverless inference provider — Inblix summary
Research

Hugging Face adds Public AI as a free, nonprofit serverless inference provider

Hugging Face Blog · Sep 17, 2025 · 2 min read

Hugging Face just made it significantly easier to run models from public-interest institutions without setting up your...

Cohere now serves its own enterprise models directly on Hugging Face Hub — Inblix summary
Research

Cohere now serves its own enterprise models directly on Hugging Face Hub

Hugging Face Blog · Apr 16, 2025 · 2 min read

Hugging Face just landed a first for its Inference Providers program: Cohere is not only joining the roster, but it's t...

Hugging Face adds Hyperbolic, Nebius, and Novita to its serverless inference army — Inblix summary
Research

Hugging Face adds Hyperbolic, Nebius, and Novita to its serverless inference army

Hugging Face Blog · Feb 18, 2025 · 2 min read

Hugging Face just dropped a trio of new serverless inference providers into its platform: Hyperbolic, Nebius AI Studio,...

Fireworks.ai lands on Hugging Face Hub with serverless access to DeepSeek-R1 — Inblix summary
Research

Fireworks.ai lands on Hugging Face Hub with serverless access to DeepSeek-R1

Hugging Face Blog · Feb 14, 2025 · 2 min read

Hugging Face just gave developers a serious speed boost by integrating Fireworks.ai directly into its Hub. Starting now...

Hugging Face Kills NVIDIA NIM Serverless Inference, Points Users to New Service — Inblix summary
Research

Hugging Face Kills NVIDIA NIM Serverless Inference, Points Users to New Service

Hugging Face Blog · Jul 29, 2024 · 2 min read

Hugging Face has quietly pulled the plug on its NVIDIA NIM API, the serverless inference service it launched with consi...