blogpost: featherless as a provider

SBrandeis · SBrandeis · commit 27d99b35f860 · 2025-06-11T11:57:33.000+02:00
diff --git a/inference-providers-feeatherless.md b/inference-providers-feeatherless.md
@@ -0,0 +1,129 @@
+---
+title: "Featherless on Hugging Face Inference Providers 🔥"
+thumbnail: /blog/assets/inference-providers-featherless/thumbnail.png
+authors:
+- user: wxgeorge
+  guest: true
+  org: featherless-ai
+- user: pohnean-recursal
+  guest: true
+  org: featherless-ai
+- user: picocreator
+  guest: true
+  org: featherless-ai
+- user: celinah
+- user: sbrandeis
+---
+
+
+![banner image](https://place-hold.it/1680x900)
+<!-- ![banner image](https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/inference-providers/featherless-banner.png) -->
+<!-- TODO: add a banner -->
+
+# Featherless on Hugging Face Inference Providers 🔥
+
+We're thrilled to share that **Featherless** is now a supported Inference Provider on HF Hub!
+Featherless joins our growing ecosystem, enhancing the breadth and capabilities of serverless inference directly on the Hub’s model pages. Inference Providers are also seamlessly integrated into our client SDKs (for both JS and Python), making it super easy to use a wide variety of models with your preferred providers.
+
+[Featherless](https://featherless.ai) supports a wide variety of text and conversational model, including the latest open-source models from DeepSeek, Meta, Google, Qwen, and much more.
+
+Check out supported models here: [supported models list](https://featherless.ai/models).
+
+We're quite excited to see what you'll build with this new provider!
+
+Read more about Inference Providers in our [documentation](https://huggingface.co/docs/inference-providers).
+
+ ## How it works
+
+### In the website UI
+
+
+1. In your user account settings, you are able to:
+- Set your own API keys for the providers you’ve signed up with. If no custom key is set, your requests will be routed through HF.
+- Order providers by preference. This applies to the widget and code snippets in the model pages.
+
+<img src="https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/inference-providers/user-settings-updated.png" alt="Inference Providers"/>
+
+
+2. As mentioned, there are two modes when calling Inference Providers: 
+- Custom key (calls go directly to the inference provider, using your own API key of the corresponding inference provider)
+- Routed by HF (in that case, you don't need a token from the provider, and the charges are applied directly to your HF account rather than the provider's account)
+
+
+<img src="https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/inference-providers/explainer.png" alt="Inference Providers"/>
+
+
+3. Model pages showcase third-party inference providers (the ones that are compatible with the current model, sorted by user preference)
+
+<img src="https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/inference-providers/model-widget-updated.png" alt="Inference Providers"/>
+
+
+### From the client SDKs
+
+#### from Python, using huggingface_hub
+
+The following example shows how to use DeepSeek-R1 using Hyperbolic as the inference provider. You can use a [Hugging Face token](https://huggingface.co/settings/tokens) for automatic routing through Hugging Face, or your own Hyperbolic API key if you have one.
+
+Install `huggingface_hub` from source (see [instructions](https://huggingface.co/docs/huggingface_hub/installation#install-from-source)). Official support will be released soon in version v0.29.0.
+
+```python
+from huggingface_hub import InferenceClient
+
+client = InferenceClient(
+    provider="featherless-ai",
+    api_key="xxxxxxxxxxxxxxxxxxxxxxxx"
+)
+
+messages = [
+    {
+        "role": "user",
+        "content": "What is the capital of France?"
+    }
+]
+
+completion = client.chat.completions.create(
+    model="deepseek-ai/DeepSeek-R1-0528", 
+    messages=messages, 
+    max_tokens=500
+)
+
+print(completion.choices[0].message)
+```
+
+#### from JS using @huggingface/inference
+
+```js
+import { HfInference } from "@huggingface/inference";
+
+const client = new HfInference("xxxxxxxxxxxxxxxxxxxxxxxx");
+
+const chatCompletion = await client.chatCompletion({
+	model: "deepseek-ai/DeepSeek-R1-0528",
+	messages: [
+		{
+			role: "user",
+			content: "What is the capital of France?"
+		}
+	],
+	provider: "featherless-ai",
+	max_tokens: 500
+});
+
+console.log(chatCompletion.choices[0].message);
+```
+
+## Billing
+
+For direct requests, i.e. when you use the key from an inference provider, you are billed by the corresponding provider. For instance, if you use a Featherless AI API key you're billed on your Featherless AI account.
+
+For routed requests, i.e. when you authenticate via the Hugging Face Hub, you'll only pay the standard provider API rates. There's no additional markup from us, we just pass through the provider costs directly. (In the future, we may establish revenue-sharing agreements with our provider partners.)
+
+**Important Note** ‼️ PRO users get $2 worth of Inference credits every month. You can use them across providers. 🔥
+
+Subscribe to the [Hugging Face PRO plan](https://hf.co/subscribe/pro) to get access to Inference credits, ZeroGPU, Spaces Dev Mode, 20x higher limits, and more.
+
+We also provide free inference with a small quota for our signed-in free users, but please upgrade to PRO if you can!
+
+## Feedback and next steps
+
+We would love to get your feedback! Here’s a Hub discussion you can use: https://huggingface.co/spaces/huggingface/HuggingDiscussions/discussions/49