Home / Companies / Together AI / Blog / Post Details
Content Deep Dive

Llama-2-7B-32K-Instruct — and fine-tuning for Llama-2 models with Together API

Blog post from Together AI

Post Details
Company
Date Published
Author
Together
Word Count
1,092
Company Posts That Month
2
Language
English
Hacker News Points
-
Post removed?
No
Summary

The Llama-2-7B-32K-Instruct model achieves state-of-the-art performance for long-context tasks such as summarization and multi-document question answering while maintaining similar performance at a shorter context length compared to the base Llama-2-7B model. The model was fine-tuned using the Together API, which allows developers to easily build custom models with less than 200 lines of Python script. The fine-tuning process involves four main steps: distilling instructions from human inputs, training the model on a mixture of data sources, testing the model in the Together Playgrounds, and deploying it via the Together Inference API. The model outperforms other baseline models including GPT-3.5-Turbo-16k, Llama-2-7b-chat, Longchat-7b-16k and Longchat-7b-v1.5-32k on long-context benchmarks, demonstrating its robustness across these tasks. The model is now available for public use with the Together API, enabling developers to build custom models with ease.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
AI Model Fine-tuning 10 670 134 68 +0%
AI Guardrails 1 87 45 22 -17%
LLM 1 3,077 361 126 +59%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.