Home / Companies / Together AI / Blog / Post Details
Content Deep Dive

Preparing for the era of 32K context: Early learnings and explorations

Blog post from Together AI

Post Details
Company
Date Published
Author
Together
Word Count
1,831
Company Posts That Month
5
Language
English
Hacker News Points
-
Post removed?
No
Summary

Together AI has released LLaMA-2-7B-32K, a 32K context model built using Position Interpolation and Together AI's data recipe and system optimizations. This model extends the original LLaMA-2 to 32K long context, achieving comparable perplexity and quality to state-of-the-art closed-source models. The power of this base model lies in its ability to be fine-tuned for targeted applications, such as multi-document question answering and summarization. To support this, Together AI has updated their inference and training stack with FlashAttention-2 and other optimizations, allowing for efficient inference and fine-tuning with 32K context. The community is encouraged to build on this work by exploring ways to extend the context length of open-source models, preparing better data for long-context tasks, and improving system support for long-context training and inference.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
AI Model Fine-tuning 6 669 87 53 +50%
LLM 2 1,935 244 98 -1%
Vector Search 2 1,161 174 75 -27%
RAG 1 144 33 19 -9%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.