← All stories
● Covered by 1 source · 1 reportMedium impact1 positive

Baseten Integrates as an Inference Provider on Hugging Face Hub

🔄 Updated 1d ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • Baseten is now an Inference Provider on Hugging Face Hub.
  • Supports conversational and text-generation tasks initially.
  • Integrates with Hugging Face client SDKs for Python and JavaScript.
  • Users can use custom API keys or route requests through Hugging Face.

Baseten Joins Hugging Face Ecosystem

Baseten has been added as a supported Inference Provider on the Hugging Face Hub. This integration allows users to access Baseten's serverless AI capabilities directly from Hugging Face's model pages and client SDKs. Baseten is an AI infrastructure platform offering serverless AI and training functionalities.

Initial Task Support and Model Access

Initially, Baseten will support conversational and text-generation tasks on Hugging Face. This enables access to various open-weight large language models (LLMs) such as Kimi K3, DeepSeek V4 Flash, and GLM-5.2. Support for additional AI tasks is planned for future rollout.

Integration with Hugging Face Tools

The integration extends to Hugging Face's client SDKs for Python (huggingface_hub >= 1.26.1) and JavaScript (@huggingface/inference). Users can configure their API keys for Baseten or route requests through Hugging Face, with charges applied to their Hugging Face account. Inference Providers are also integrated into Agent Harnesses like Pi and OpenCode, allowing direct use of Baseten-hosted models.

Impact on Developers

This partnership provides developers with more options for deploying and interacting with AI models. By integrating Baseten, Hugging Face expands its ecosystem of serverless inference providers, simplifying the process for developers to incorporate a wider range of AI capabilities into their applications with minimal setup.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~7 min · 6 stories · Aug 15

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Primary sources

GitHub huggingface/blog

Reporting from

Baseten is now a supported Inference Provider on the Hugging Face Hub, allowing developers to use Baseten's serverless AI platform for conversational and text-generation tasks directly from Hugging Face model pages and SDKs. This integration expands the options for deploying and utilizing open-weight large language models within the Hugging Face ecosystem.