Today's AI/ML headlines are brought to you by ThreatPerspective

Digital Event Horizon

Baseten Joins Hugging Face Inference Providers as a Supported Inference Provider



Baseten, an AI infrastructure platform, has been integrated as a supported inference provider on Hugging Face, enhancing serverless inference capabilities and providing developers with access to popular LLMs. This partnership aims to streamline AI development and deployment, offering unparalleled flexibility and scalability.

  • Baseten's AI infrastructure platform has been integrated into Hugging Face Hub as a supported inference provider.
  • The integration aims to enhance serverless inference capabilities on the Hugging Face Hub's model pages.
  • Baseten supports various models, including large language models, text-to-speech, and other cutting-edge technologies.
  • Users can now seamlessly leverage Baseten and Hugging Face capabilities, providing flexibility and scalability for AI projects.
  • User-friendly features include setting API keys, ordering providers by preference, and custom key modes.
  • The integration brings several benefits to PRO users, including $2 worth of Inference credits every month and access to additional features.


  • Hugging Face, a prominent artificial intelligence (AI) research organization, has officially welcomed Baseten, an AI infrastructure platform, to its esteemed list of supported inference providers. This significant integration is aimed at enhancing the breadth and capabilities of serverless inference directly on the Hugging Face Hub's model pages.

    Baseten's AI infrastructure platform offers a comprehensive suite of tools for developers to integrate a wide range of AI capabilities into their applications with minimal setup. The platform supports various models, including large language models (LLMs), text-to-speech, and other cutting-edge technologies. As part of its initial integration, Baseten is launching support for conversational and text-generation tasks on Hugging Face, enabling access to popular open-weight LLMs such as Kimi K3, DeepSeek V4 Flash, GLM-5.2, and many more.

    The new integration allows users to seamlessly leverage the capabilities of both Baseten and Hugging Face, providing unparalleled flexibility and scalability for their AI projects. With this partnership, developers can utilize a wide range of models and inference providers through Baseten's API or client SDKs, making it easier to adopt serverless inference on the Hugging Face Hub.

    Furthermore, users can set their own API keys for supported providers, order them by preference, and choose between custom key mode (direct calls to the provider) and routed mode (calls authenticated via the Hugging Face Hub). This flexibility enables developers to optimize their workflow and manage costs effectively.

    The support for Baseten as a trusted inference provider comes with several benefits. For instance, PRO users will receive $2 worth of Inference credits every month, which can be used across providers. Additionally, users can upgrade to the Hugging Face PRO plan to gain access to more features, including ZeroGPU, Spaces Dev Mode, and higher limits.

    The addition of Baseten to the list of supported inference providers further solidifies Hugging Face's commitment to providing a comprehensive platform for developers to build and deploy AI models. As the AI landscape continues to evolve, this integration is poised to bring about significant advancements in the field, empowering users with cutting-edge tools and infrastructure.



    Related Information:
  • https://www.digitaleventhorizon.com/articles/Baseten-Joins-Hugging-Face-Inference-Providers-as-a-Supported-Inference-Provider-deh.shtml

  • https://huggingface.co/blog/baseten


  • Published: Thu Aug 6 09:56:34 2026 by llama3.2 3B Q4_K_M











    © Digital Event Horizon . All rights reserved.

    Privacy | Terms of Use | Contact Us