Home / Companies / LllamaIndex / Blog / Post Details
Content Deep Dive

Context

Blog post from LllamaIndex

Post Details
Company
Date Published
Author
Jerry Liu
Word Count
1,230
Company Posts That Month
6
Language
English
Hacker News Points
-
Post removed?
No
Summary

LlamaIndex has introduced a new feature that allows users to fine-tune a linear adapter on top of query embeddings from any model, such as sentence_transformers and OpenAI, to optimize retrieval over specific data and queries without needing to re-embed documents. This method, inspired by Vespa's concept of "freezing" document embeddings, involves the use of EmbeddingAdapterFinetuneEngine, which transforms only the query embeddings while keeping document embeddings fixed. The fine-tuning process is straightforward, can be performed on various embedding models, and is shown to result in modest performance improvements in retrieval metrics like hit-rate and mean reciprocal rank. Despite the modest gains compared to directly fine-tuning sentence_transformers, this approach offers a cost-effective and flexible solution for users looking to enhance retrieval performance without extensive reprocessing of document embeddings.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Vector Search 43 1,500 202 67 -14%
AI Model Fine-tuning 18 498 94 48 -24%
RAG 2 466 92 33 +83%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.