Home / Companies / Voyage AI / Blog / Post Details
Content Deep Dive

Domain-Specific Embeddings and Retrieval: Legal Edition (voyage-law-2)

Blog post from Voyage AI

Post Details
Company
Date Published
Author
Voyage AI
Word Count
1,180
Company Posts That Month
1
Language
English
Hacker News Points
-
Post removed?
No
Summary

Voyage-law-2 is a newly released domain-specific embedding model optimized for legal document retrieval, significantly outperforming general-purpose models like OpenAI v3 large, particularly in legal contexts. Trained on a vast dataset of legal documents, it features a 16K-context length, excelling in long-context retrieval. On eight legal retrieval datasets, voyage-law-2 led in seven, including notable performance on LeCaRDv2, LegalQuAD, and GerDaLIR with over 10% improvement in comparison to competitors. The model also demonstrates strong cross-domain capabilities, having been trained on various domains to enhance its applicability outside the legal field. It surpasses OpenAI v3 large in retrieval tasks across 34 datasets and eight categories, including technical documentation, finance, and medicine, indicating its robust adaptability and effectiveness in various contexts.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Vector Search 13 2,613 257 91 +44%
RAG 3 1,795 223 72 +55%
AI Model Fine-tuning 1 742 135 73 +71%
LLM 1 3,398 379 136 +44%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.