# Groq on Hugging Face Inference Providers 🔥

DevFeed: [Groq on Hugging Face Inference Providers 🔥](<https://devfeed.tech/articles/groq-on-hugging-face-inference-providers-7281.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/inference-providers-groq>)

Author: Ben Ankiel; Hatice Ozen; Célina Hanouti; Lucain Pouget; Simon Brandeis

Published: 2025-06-16T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [groq](<https://devfeed.tech/topics/groq.md>), [inference-providers](<https://devfeed.tech/topics/inference-providers.md>), [hugging face](<https://devfeed.tech/topics/hugging-face.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [model-deployment](<https://devfeed.tech/topics/model-deployment.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [SDKs](<https://devfeed.tech/topics/sdks.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [api](<https://devfeed.tech/tags/api.md>), [api-keys](<https://devfeed.tech/tags/api-keys.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [groq](<https://devfeed.tech/tags/groq.md>), [hub](<https://devfeed.tech/tags/hub.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [inference](<https://devfeed.tech/tags/inference.md>), [inference-providers](<https://devfeed.tech/tags/inference-providers.md>), [large-language-models-llms](<https://devfeed.tech/tags/large-language-models-llms.md>), [latency](<https://devfeed.tech/tags/latency.md>), [llms](<https://devfeed.tech/tags/llms.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [partnerships](<https://devfeed.tech/tags/partnerships.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [sdks](<https://devfeed.tech/tags/sdks.md>)

## AI overview

Groq is now available as an Inference Provider on the Hugging Face Hub, including model pages and Hugging Face client SDKs for JavaScript and Python. The article describes Groq's LPU technology, its low-latency and high-throughput inference for LLMs, support for open models such as Meta's Llama 4 and Qwen's QWQ-32B, API access, and custom-key or Hugging Face-routed usage options.

## Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.