Baseten on Hugging Face Inference Providers 🔥
Blog post from Hugging Face
Hugging Face has added Baseten as a supported Inference Provider, expanding serverless inference options available through Hub model pages, Python and JavaScript SDKs, and several agent harnesses. Baseten’s initial integration supports conversational and text-generation workloads using open-weight models including Kimi K3, DeepSeek V4 Flash, and GLM-5.2, with additional task types planned. Users can configure provider preferences and either supply a personal Baseten API key for direct billing or use a Hugging Face token for routed requests billed at standard provider rates without Hugging Face markup. The integration is available in huggingface_hub version 1.26.1 or later and the @huggingface/inference JavaScript package, while Hugging Face PRO subscribers receive $2 in monthly inference credits usable across providers.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 4 | 5,068 | 1,020 | 229 | -34% |
| Serverless | 2 | 783 | 217 | 99 | +1% |
| OpenClaw | 1 | 184 | 39 | 19 | -39% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.