# inference-providers

Hugging Face Inference Providers is a developer-facing platform that gives access to machine-learning models through multiple inference providers and consistent APIs and SDKs.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## How to Move an AI App from a Frontier Model to Open Weights

DevFeed: [How to Move an AI App from a Frontier Model to Open Weights](<https://devfeed.tech/articles/how-to-move-an-ai-app-from-a-frontier-model-to-open-weights-32182.md>)

Original publisher: [Read original article](<https://spin.atomicobject.com/ai-app-frontier-open-weights/>)

Author: Gus Schissler

Published: 2026-09-10T12:00:43Z

Content type: tutorial

Language: en

Sources: [Atomic Object](<https://devfeed.tech/sources/atomic-object.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>), [Benchmark](<https://devfeed.tech/topics/benchmark.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [codex](<https://devfeed.tech/topics/codex.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-for-developers](<https://devfeed.tech/tags/ai-for-developers.md>), [benchmark](<https://devfeed.tech/tags/benchmark.md>), [codex](<https://devfeed.tech/tags/codex.md>), [how-to](<https://devfeed.tech/tags/how-to.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [llm](<https://devfeed.tech/tags/llm.md>), [prototyping](<https://devfeed.tech/tags/prototyping.md>)

### AI overview

This tutorial describes moving an AI document-processing prototype from a frontier model to cheaper open-weight models through hosted inference providers. It explains how to define a trusted gold set from real project outputs and use workflow-specific evaluation to identify where lower-cost models fail.

### Source excerpt

At Atomic Object, we're encouraged to prototype and dogfood internal projects. Over the last four months, I've been building a prototype to help solve a part of my job that I dislike. The first version used an LLM to extract and classify information from documents, then connect related pieces so humans and agents could retrieve [...] The post How to Move an AI App from a Frontier Model to Open Weights appeared first on Atomic Spin.

## NVIDIA to Acquire Hugging Face for $12.93B, Pledges the Platform Stays Open and Hardware Neutral

DevFeed: [NVIDIA to Acquire Hugging Face for $12.93B, Pledges the Platform Stays Open and Hardware Neutral](<https://devfeed.tech/articles/nvidia-to-acquire-hugging-face-for-12-93b-pledges-the-platform-stays-open-and-hardware-neutral-12368.md>)

Original publisher: [Read original article](<https://www.storagereview.com/news/nvidia-to-acquire-hugging-face-for-12-93b-pledges-the-platform-stays-open-and-hardware-neutral>)

Author: Harold Fritts

Published: 2026-09-04T18:01:12Z

Content type: news

Language: en

Sources: [StorageReview.com](<https://devfeed.tech/sources/storagereview-com.md>)

Topics: [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [Nvidia](<https://devfeed.tech/topics/nvidia.md>), [AI Development](<https://devfeed.tech/topics/ai-development.md>), [AI Infrastructure](<https://devfeed.tech/topics/ai-infrastructure.md>), [Application Development](<https://devfeed.tech/topics/application-development.md>), [datasets](<https://devfeed.tech/topics/datasets.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [Deployment](<https://devfeed.tech/topics/deployment.md>), [model-deployment](<https://devfeed.tech/topics/model-deployment.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-infrastructure](<https://devfeed.tech/tags/ai-infrastructure.md>), [application-development](<https://devfeed.tech/tags/application-development.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [creators](<https://devfeed.tech/tags/creators.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [deployment](<https://devfeed.tech/tags/deployment.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [models](<https://devfeed.tech/tags/models.md>), [multi-cloud](<https://devfeed.tech/tags/multi-cloud.md>), [nvidia](<https://devfeed.tech/tags/nvidia.md>), [platforms](<https://devfeed.tech/tags/platforms.md>)

### AI overview

NVIDIA has agreed to acquire Hugging Face for $12.93 billion, with plans to expand its infrastructure and AI development capabilities. Hugging Face is expected to retain its brand and operate as an open, hardware-neutral platform supporting models, datasets, applications, multiple clouds, accelerators, and inference providers.

### Source excerpt

NVIDIA has agreed to acquire Hugging Face for $12.93 billion, a transaction that would extend the company's position from accelerated compute and AI infrastructure into one of the industry's most widely used platforms for open models, datasets, and application development. In an announcement published on the NVIDIA website, CEO Jensen Huang said the company plans The post NVIDIA to Acquire Hugging Face for $12.93B, Pledges the Platform Stays Open and Hardware Neutral appeared first on StorageReview.com.

## Access and share AI Gateway leaderboard data

DevFeed: [Access and share AI Gateway leaderboard data](<https://devfeed.tech/articles/access-and-share-ai-gateway-leaderboard-data-1036.md>)

Original publisher: [Read original article](<https://vercel.com/changelog/open-data-and-shareable-charts-for-ai-gateway-leaderboards>)

Author: Jerilyn Zheng

Published: 2026-07-14T00:00:00Z

Content type: release

Language: en

Sources: [Vercel News](<https://devfeed.tech/sources/vercel-news.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [API](<https://devfeed.tech/topics/api.md>), [CSV](<https://devfeed.tech/topics/csv.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [data](<https://devfeed.tech/topics/data.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [api](<https://devfeed.tech/tags/api.md>), [data](<https://devfeed.tech/tags/data.md>), [images](<https://devfeed.tech/tags/images.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [leaderboard](<https://devfeed.tech/tags/leaderboard.md>), [metrics](<https://devfeed.tech/tags/metrics.md>), [model](<https://devfeed.tech/tags/model.md>), [open](<https://devfeed.tech/tags/open.md>), [ranking](<https://devfeed.tech/tags/ranking.md>), [time](<https://devfeed.tech/tags/time.md>), [tokens](<https://devfeed.tech/tags/tokens.md>), [videos](<https://devfeed.tech/tags/videos.md>)

### AI overview

Vercel has opened the data behind the AI Gateway leaderboards under the CC BY 4.0 license. The data can be downloaded as CSV or queried through an export API, while charts can be shared as branded PNG images. The leaderboards track daily production usage across models, labs, apps, and inference providers.

### Source excerpt

We are making the data behind the AI Gateway leaderboards open under the CC BY 4.0 license. You can now download or query the data through the leaderboard-export API endpoint and render any chart as a shareable image. The AI Gateway leaderboards show how AI is used in production, ranking traffic for models, labs, apps, and providers. Data is aggregated daily across trillions of tokens, so you can see what gets adopted and how that changes over time. For deeper analysis, see the July AI Gateway Production Index. What's ranked There are four leaderboards, each with its own metrics: Leaderboard Ranks Metrics Models Individual models Requests, token volume, spend, images or videos generated Labs Model labs Requests, token volume, spend, images or videos generated Apps Opted-in apps built on the AI Gateway Token volume, spend Providers Inference providers Token volume, spend Models and labs can be filtered by modality (text, image, video) and show a daily percentage share over time; apps and providers are aggregated across all modalities and show a ranked top list. Open data The data behind the leaderboards is open, published under Creative Commons Attribution 4.0 (CC BY 4.0). You are free to use, share, and adapt it, including commercially, as long as you give credit, link to the license, and indicate if changes were made. Every chart and ranked list has a download button that exports the current view as a CSV. For programmatic access, use the export endpoint, which returns the same data and is cached for 24 hours: For models and labs, each row is one entity's daily share of a single metric. One response includes rows for requests, tokens, spend, imageCount, and videoCount, so filter on the metric field to pull out the series you want. Share a chart Every chart has a share button that turns the current view into an image. Pick an aspect ratio (landscape, square, or portrait), then download it as a PNG or copy it to your clipboard. The image includes the legend, title, a

## DeepInfra on Hugging Face Inference Providers 🔥

DevFeed: [DeepInfra on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/deepinfra-on-hugging-face-inference-providers-7279.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-deepinfra>)

Author: Aray Sultanbekova; Shang-Pin; Utemuratov; Yessen K; Oguz Vuruskaner; Célina Hanouti; Simon Brandeis; Lucain Pouget

Published: 2026-04-29T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [deepinfra](<https://devfeed.tech/topics/deepinfra.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [AI Inference](<https://devfeed.tech/topics/ai-inference.md>), [SDKs](<https://devfeed.tech/topics/sdks.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [text-generation](<https://devfeed.tech/topics/text-generation.md>), [text-to-image](<https://devfeed.tech/topics/text-to-image.md>)

Tags: [ai-inference](<https://devfeed.tech/tags/ai-inference.md>), [api](<https://devfeed.tech/tags/api.md>), [api-keys](<https://devfeed.tech/tags/api-keys.md>), [applications](<https://devfeed.tech/tags/applications.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [code](<https://devfeed.tech/tags/code.md>), [deepinfra](<https://devfeed.tech/tags/deepinfra.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [javascript](<https://devfeed.tech/tags/javascript.md>), [llms](<https://devfeed.tech/tags/llms.md>), [open](<https://devfeed.tech/tags/open.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [pricing](<https://devfeed.tech/tags/pricing.md>), [python](<https://devfeed.tech/tags/python.md>), [sdks](<https://devfeed.tech/tags/sdks.md>), [text-generation](<https://devfeed.tech/tags/text-generation.md>), [text-to-image](<https://devfeed.tech/tags/text-to-image.md>)

### AI overview

DeepInfra is now a supported Inference Provider on the Hugging Face Hub, with serverless access to more than 100 models and integration with Hugging Face's JavaScript and Python SDKs. The article describes provider selection, API-key and routed-by-Hugging-Face modes, supported model tasks, and initial access to conversational and text-generation models.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Liberate your OpenClaw

DevFeed: [Liberate your OpenClaw](<https://devfeed.tech/articles/liberate-your-openclaw-7331.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/liberate-your-openclaw>)

Author: Clem 🤗; ben burtenshaw; Pedro Cuenca; Jeff Boudier; merve; Niels Rogge; Victor Mustar; Mishig ᠮᠢᠰᠾᠢᠭ

Published: 2026-03-27T00:00:00Z

Content type: tutorial

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [OpenClaw](<https://devfeed.tech/topics/openclaw.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [llama.cpp](<https://devfeed.tech/topics/llama-cpp.md>)

Tags: [agents](<https://devfeed.tech/tags/agents.md>), [guide](<https://devfeed.tech/tags/guide.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [llama-cpp](<https://devfeed.tech/tags/llama-cpp.md>), [local-ai](<https://devfeed.tech/tags/local-ai.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [openclaw](<https://devfeed.tech/tags/openclaw.md>), [privacy](<https://devfeed.tech/tags/privacy.md>), [server](<https://devfeed.tech/tags/server.md>)

### AI overview

This tutorial explains how to restore OpenClaw agents using open models through Hugging Face Inference Providers or by running models locally with llama.cpp. It compares hosted and local approaches, covering privacy, cost, hardware, model selection, configuration, and local server setup.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Open Responses: What you need to know

DevFeed: [Open Responses: What you need to know](<https://devfeed.tech/articles/open-responses-what-you-need-to-know-7426.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/open-responses>)

Author: shaun smith; ben burtenshaw; merve; Pedro Cuenca

Published: 2026-01-15T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [open-responses](<https://devfeed.tech/topics/open-responses.md>), [API](<https://devfeed.tech/topics/api.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [JSON](<https://devfeed.tech/topics/json.md>), [OpenAI](<https://devfeed.tech/topics/openai.md>)

Tags: [agentic](<https://devfeed.tech/tags/agentic.md>), [agents](<https://devfeed.tech/tags/agents.md>), [api](<https://devfeed.tech/tags/api.md>), [apis](<https://devfeed.tech/tags/apis.md>), [autonomous](<https://devfeed.tech/tags/autonomous.md>), [community](<https://devfeed.tech/tags/community.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [json](<https://devfeed.tech/tags/json.md>), [mcp](<https://devfeed.tech/tags/mcp.md>), [open-responses](<https://devfeed.tech/tags/open-responses.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [openai](<https://devfeed.tech/tags/openai.md>), [responses](<https://devfeed.tech/tags/responses.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [tool](<https://devfeed.tech/tags/tool.md>), [workflow](<https://devfeed.tech/tags/workflow.md>)

### AI overview

Open Responses is presented as an open inference standard for agentic workloads. The article explains how it extends and open-sources the Responses API, supports interoperable routing across inference providers, and introduces standardized configuration, semantic streaming events, extensibility, and optional reasoning fields.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## OVHcloud on Hugging Face Inference Providers 🔥

DevFeed: [OVHcloud on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/ovhcloud-on-hugging-face-inference-providers-7027.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/OVHcloud/inference-providers-ovhcloud>)

Author: Gilles Closset; Fabien Ric; Elias Tourneux

Published: 2025-11-24T16:08:47Z

Content type: release

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [Serverless](<https://devfeed.tech/topics/serverless.md>), [SDKs](<https://devfeed.tech/topics/sdks.md>), [AI Models](<https://devfeed.tech/topics/ai-models.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [ai-models](<https://devfeed.tech/tags/ai-models.md>), [api](<https://devfeed.tech/tags/api.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [latency](<https://devfeed.tech/tags/latency.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [production](<https://devfeed.tech/tags/production.md>), [sdks](<https://devfeed.tech/tags/sdks.md>), [serverless](<https://devfeed.tech/tags/serverless.md>), [text-generation](<https://devfeed.tech/tags/text-generation.md>), [tokens](<https://devfeed.tech/tags/tokens.md>)

### AI overview

Hugging Face announces OVHcloud as a supported Inference Provider on the Hugging Face Hub. The serverless service provides managed access to open-weight and frontier AI models through API calls, with SDK integration, European infrastructure, pay-per-token pricing, low latency, multimodal capabilities, structured outputs, function calling, text generation, and embeddings.

### Source excerpt

We're thrilled to share that OVHcloud is now a supported Inference Provider on the Hugging Face Hub! OVHcloud joins our growing ecosystem, enhancing the breadth and capabilities of serverless inference directly on the Hub's model pages. Inference Providers are also seamlessly integrated into our client SDKs (for both JS and Python), making it super easy to use a wide variety of models with your preferred providers.

## Unlock the power of images with AI Sheets

DevFeed: [Unlock the power of images with AI Sheets](<https://devfeed.tech/articles/unlock-the-power-of-images-with-ai-sheets-7080.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/aisheets-unlock-images>)

Author: Ame Vi; Daniel Vila; Francisco Aranda; Damián Pumar; Leandro von Werra; Thomas Wolf

Published: 2025-10-21T00:00:00Z

Content type: release

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [AI Development](<https://devfeed.tech/topics/ai-development.md>), [AI Models](<https://devfeed.tech/topics/ai-models.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [datasets](<https://devfeed.tech/topics/datasets.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [CSV](<https://devfeed.tech/topics/csv.md>), [parquet](<https://devfeed.tech/topics/parquet.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-models](<https://devfeed.tech/tags/ai-models.md>), [code](<https://devfeed.tech/tags/code.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [parquet](<https://devfeed.tech/tags/parquet.md>), [release](<https://devfeed.tech/tags/release.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Hugging Face has released a major update to AI Sheets, an open-source spreadsheet tool that uses AI models to build, transform, and enrich datasets without code. The update adds vision support for analyzing images, extracting structured information, generating visuals, and editing images, with access to thousands of open models through Inference Providers.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Scaleway on Hugging Face Inference Providers 🔥

DevFeed: [Scaleway on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/scaleway-on-hugging-face-inference-providers-7284.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-scaleway>)

Author: Guillaume Noale; Franck Pagny; Fred Bardolle; Guillaume Calmettes; Constance Morales; Célina Hanouti; Julien Chaumond; Simon Brandeis; Lucain Pouget

Published: 2025-09-19T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [scaleway](<https://devfeed.tech/topics/scaleway.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [Serverless](<https://devfeed.tech/topics/serverless.md>), [gpt-oss](<https://devfeed.tech/topics/gpt-oss.md>), [SDKs](<https://devfeed.tech/topics/sdks.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [Latency](<https://devfeed.tech/topics/latency.md>), [Sovereign AI](<https://devfeed.tech/topics/sovereign-ai.md>)

Tags: [api-keys](<https://devfeed.tech/tags/api-keys.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [function-calling](<https://devfeed.tech/tags/function-calling.md>), [gemma](<https://devfeed.tech/tags/gemma.md>), [gpt-oss](<https://devfeed.tech/tags/gpt-oss.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [llms](<https://devfeed.tech/tags/llms.md>), [low-latency](<https://devfeed.tech/tags/low-latency.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [open](<https://devfeed.tech/tags/open.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [providers](<https://devfeed.tech/tags/providers.md>), [qwen3](<https://devfeed.tech/tags/qwen3.md>), [scaleway](<https://devfeed.tech/tags/scaleway.md>), [sdks](<https://devfeed.tech/tags/sdks.md>), [serverless](<https://devfeed.tech/tags/serverless.md>), [text-generation](<https://devfeed.tech/tags/text-generation.md>)

### AI overview

Hugging Face announces Scaleway as a supported Inference Provider on the Hugging Face Hub. The integration provides serverless access to open-weight and frontier AI models through model pages, client SDKs, and APIs, with European data centers, pay-per-token pricing, low latency, and production-oriented features.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Public AI on Hugging Face Inference Providers 🔥

DevFeed: [Public AI on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/public-ai-on-hugging-face-inference-providers-7283.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-publicai>)

Author: Joseph Low; Joshua Tan; Célina Hanouti; Julien Chaumond; Simon Brandeis; Lucain Pouget

Published: 2025-09-17T00:00:00Z

Content type: release

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [ai-inference](<https://devfeed.tech/tags/ai-inference.md>), [api](<https://devfeed.tech/tags/api.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [llms](<https://devfeed.tech/tags/llms.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [public-ai](<https://devfeed.tech/tags/public-ai.md>)

### AI overview

Hugging Face announces Public AI as a supported Inference Provider on the Hub. Public AI offers access to public and sovereign models through distributed vLLM-based infrastructure and OpenAI-compatible APIs.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Welcome GPT OSS, the new open-source model family from OpenAI!

DevFeed: [Welcome GPT OSS, the new open-source model family from OpenAI!](<https://devfeed.tech/articles/welcome-gpt-oss-the-new-open-source-model-family-from-openai-7567.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/welcome-openai-gpt-oss>)

Author: Vaibhav Srivastav; Pedro Cuenca; Lewis Tunstall; Clem 🤗; Matthew Carrigan; Clémentine Fourrier; Célina Hanouti; Lucain Pouget; Marc Sun; Simon Pagezy

Published: 2025-08-05T00:00:00Z

Content type: release

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [gpt-oss](<https://devfeed.tech/topics/gpt-oss.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [OpenAI](<https://devfeed.tech/topics/openai.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [llama.cpp](<https://devfeed.tech/topics/llama-cpp.md>), [moe](<https://devfeed.tech/topics/moe.md>), [Ollama](<https://devfeed.tech/topics/ollama.md>), [Python](<https://devfeed.tech/topics/python.md>), [quantization](<https://devfeed.tech/topics/quantization.md>), [Transformers](<https://devfeed.tech/topics/transformers.md>), [vllm](<https://devfeed.tech/topics/vllm.md>)

Tags: [api](<https://devfeed.tech/tags/api.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [community](<https://devfeed.tech/tags/community.md>), [gpt](<https://devfeed.tech/tags/gpt.md>), [gpt-oss](<https://devfeed.tech/tags/gpt-oss.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [javascript](<https://devfeed.tech/tags/javascript.md>), [llama-cpp](<https://devfeed.tech/tags/llama-cpp.md>), [llm](<https://devfeed.tech/tags/llm.md>), [moe](<https://devfeed.tech/tags/moe.md>), [ollama](<https://devfeed.tech/tags/ollama.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [openai](<https://devfeed.tech/tags/openai.md>), [python](<https://devfeed.tech/tags/python.md>), [quantization](<https://devfeed.tech/tags/quantization.md>), [transformers](<https://devfeed.tech/tags/transformers.md>), [vllm](<https://devfeed.tech/tags/vllm.md>)

### AI overview

Hugging Face welcomes OpenAI's gpt-oss open-source model family. The article describes the models' Apache 2.0 licensing, local deployment options, reasoning and tool-use capabilities, MoE architecture, quantization, supported inference implementations, and access through Inference Providers and the Responses API.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Groq on Hugging Face Inference Providers 🔥

DevFeed: [Groq on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/groq-on-hugging-face-inference-providers-7281.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-groq>)

Author: Ben Ankiel; Hatice Ozen; Célina Hanouti; Lucain Pouget; Simon Brandeis

Published: 2025-06-16T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [groq](<https://devfeed.tech/topics/groq.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [model-deployment](<https://devfeed.tech/topics/model-deployment.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [SDKs](<https://devfeed.tech/topics/sdks.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [api](<https://devfeed.tech/tags/api.md>), [api-keys](<https://devfeed.tech/tags/api-keys.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [groq](<https://devfeed.tech/tags/groq.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [large-language-models-llms](<https://devfeed.tech/tags/large-language-models-llms.md>), [latency](<https://devfeed.tech/tags/latency.md>), [llms](<https://devfeed.tech/tags/llms.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [sdks](<https://devfeed.tech/tags/sdks.md>)

### AI overview

Groq is now available as an Inference Provider on the Hugging Face Hub, including model pages and Hugging Face client SDKs for JavaScript and Python. The article describes Groq's LPU technology, its low-latency and high-throughput inference for LLMs, support for open models such as Meta's Llama 4 and Qwen's QWQ-32B, API access, and custom-key or Hugging Face-routed usage options.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Featherless AI on Hugging Face Inference Providers 🔥

DevFeed: [Featherless AI on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/featherless-ai-on-hugging-face-inference-providers-7280.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-featherless>)

Author: Wesley George; Poh Nean; Eugene Cheah; Célina Hanouti; Lucain Pouget; Simon Brandeis

Published: 2025-06-12T00:00:00Z

Content type: release

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [AI Inference](<https://devfeed.tech/topics/ai-inference.md>), [SDKs](<https://devfeed.tech/topics/sdks.md>), [Serverless](<https://devfeed.tech/topics/serverless.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [Orchestration](<https://devfeed.tech/topics/orchestration.md>)

Tags: [ai-inference](<https://devfeed.tech/tags/ai-inference.md>), [api-keys](<https://devfeed.tech/tags/api-keys.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [featherless](<https://devfeed.tech/tags/featherless.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [llms](<https://devfeed.tech/tags/llms.md>), [orchestration](<https://devfeed.tech/tags/orchestration.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [recursal](<https://devfeed.tech/tags/recursal.md>), [sdks](<https://devfeed.tech/tags/sdks.md>), [serverless](<https://devfeed.tech/tags/serverless.md>)

### AI overview

Hugging Face announces Featherless AI as a supported Inference Provider on the Hugging Face Hub. The article explains its serverless inference offering, provider routing options, API key handling, and use through the Hugging Face client SDKs.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Cohere on Hugging Face Inference Providers 🔥

DevFeed: [Cohere on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/cohere-on-hugging-face-inference-providers-7278.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-cohere>)

Author: Vaibhav Srivastav; ben burtenshaw; merve; Célina Hanouti; Alejandro Rodriguez; Julien Chaumond; Simon Brandeis

Published: 2025-04-16T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [cohere](<https://devfeed.tech/topics/cohere.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [Retrieval Augmented Generation (RAG)](<https://devfeed.tech/topics/retrieval-augmented-generation-rag.md>), [multimodal](<https://devfeed.tech/topics/multimodal.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [cohere](<https://devfeed.tech/tags/cohere.md>), [embeddings](<https://devfeed.tech/tags/embeddings.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [generative-ai](<https://devfeed.tech/tags/generative-ai.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [llms](<https://devfeed.tech/tags/llms.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [open](<https://devfeed.tech/tags/open.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [rag](<https://devfeed.tech/tags/rag.md>), [ranking](<https://devfeed.tech/tags/ranking.md>)

### AI overview

Cohere is now supported as an Inference Provider on the Hugging Face Hub, allowing serverless inference for a selection of Cohere and Cohere Labs models. The article highlights models for enterprise AI, multilingual applications, retrieval-augmented generation, tool use, low-cost or low-latency workloads, and vision-language tasks.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Introducing Three New Serverless Inference Providers: Hyperbolic, Nebius AI Studio, and Novita 🔥

DevFeed: [Introducing Three New Serverless Inference Providers: Hyperbolic, Nebius AI Studio, and Novita 🔥](<https://devfeed.tech/articles/introducing-three-new-serverless-inference-providers-hyperbolic-nebius-ai-studio-and-novita-7282.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-nebius-novita-hyperbolic>)

Author: Julien Chaumond; Bertrand Chevrier; Vaibhav Srivastav; Simon Brandeis; Albert Abdulmanov; Viktor Hu; Connor Chevli

Published: 2025-02-18T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [model-deployment](<https://devfeed.tech/topics/model-deployment.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>)

Tags: [announcement](<https://devfeed.tech/tags/announcement.md>), [api](<https://devfeed.tech/tags/api.md>), [api-keys](<https://devfeed.tech/tags/api-keys.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [image](<https://devfeed.tech/tags/image.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>)

### AI overview

Hugging Face introduces three new serverless inference providers: Hyperbolic, Nebius AI Studio, and Novita. The providers add access to models such as DeepSeek-R1 and Flux.1, with support for custom provider keys or routing through Hugging Face.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Welcome to Inference Providers on the Hub 🔥

DevFeed: [Welcome to Inference Providers on the Hub 🔥](<https://devfeed.tech/articles/welcome-to-inference-providers-on-the-hub-7277.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers>)

Author: Burkay Gur; Zeke Sikelianos; Anton McGonnell; Hassan El Mghari; Simon Brandeis; Bertrand Chevrier; Julien Chaumond

Published: 2025-01-28T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [Open Source Models & Datasets](<https://devfeed.tech/topics/open-source-models-datasets.md>)

Tags: [announcement](<https://devfeed.tech/tags/announcement.md>), [api-keys](<https://devfeed.tech/tags/api-keys.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [serverless](<https://devfeed.tech/tags/serverless.md>), [token](<https://devfeed.tech/tags/token.md>)

### AI overview

Hugging Face introduces Inference Providers on the Hub, offering unified access to serverless model inference through partner providers. Users can configure provider API keys and preferences, while requests may either go directly to a provider or be routed through Hugging Face.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.