Home / Companies / Unstructured / Blog / Post Details
Content Deep Dive

RAG vs. Long-Context Models. Do we still need RAG?

Blog post from Unstructured

Post Details
Company
Date Published
Author
Maria Khalusova
Word Count
1,660
Company Posts That Month
12
Language
English
Hacker News Points
-
Post removed?
No
Summary

Retrieval augmented generation (RAG) is a technique that enhances large language models (LLMs) by augmenting their text generation with relevant information from external knowledge bases, addressing limitations like hallucinations common in models trained solely on publicly available data. While the expansion of context windows in LLMs, such as Gemini 1.5 Pro's 2 million-token capacity and the possibility of models with infinite context windows, offers potential advantages, RAG remains crucial for its efficiency, scalability, and cost-effectiveness. It provides transparency and accountability by allowing LLMs to trace information back to its source, which is critical in sectors like finance, healthcare, and law. RAG also facilitates role-based access control by retrieving only necessary information for specific queries, further enhancing data security. Despite the promise of long-context models, RAG's ability to efficiently retrieve and manage diverse data sources, coupled with its computational efficiency and transparency, ensures its continued relevance, even in a future where infinite context models might exist.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
RAG 34 2,243 291 87 +14%
LLM 24 3,988 514 165 -1%
AI Agents 1 515 134 62 -21%
AI Guardrails 1 292 74 39 +93%
Data Pipeline 1 747 237 70 -48%
Observability 1 1,969 341 98 +10%
Vector Search 1 4,713 314 102 +27%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.