Introducing Functions and Model Inference on Zilliz Cloud: Automatic Embedding and Reranking with Hosted Models
Blog post from Zilliz
Zilliz Cloud introduces Functions and Inference Services, designed to automate the embedding and reranking processes in AI search pipelines, which traditionally require manual embedding generation and separate reranking services. Now available in Public Preview for third-party models and Private Preview for Zilliz Hosted Models, these services allow users to insert raw text and conduct searches with natural language, while Zilliz Cloud manages embedding generation, vector storage, and result reranking. Functions are categorized into Pre-search Functions, which convert text into searchable representations, and Post-search Functions, which refine and reorder results. The Inference Services power these model-based Functions by calling models from third-party providers or utilizing Zilliz's GPU infrastructure for Hosted Models, ensuring data privacy and reducing latency. This integration simplifies the search workflow by eliminating the need for client-side embedding code and guaranteeing model consistency, ultimately enhancing search quality and operational efficiency for scalable applications.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Vector Search | 47 | 3,215 | 679 | 175 | +33% |
| Data Pipeline | 2 | 1,290 | 393 | 99 | +171% |
| LLM | 1 | 7,531 | 1,250 | 268 | +26% |
| Observability | 1 | 4,660 | 984 | 209 | +14% |
| Secrets Management | 1 | 1,946 | 398 | 127 | +28% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.