Home / Companies / Refuel / Blog / Post Details
Content Deep Dive

Announcing Refuel LLM-2

Blog post from Refuel

Post Details
Company
Date Published
Author
Refuel Team
Word Count
1,393
Company Posts That Month
1
Language
English
Hacker News Points
-
Post removed?
No
Summary

RefuelLLM-2 and RefuelLLM-2-small are new large language models specifically designed for data labeling, enrichment, and cleaning, demonstrating superior performance compared to state-of-the-art models like GPT-4-Turbo and Claude-3-Opus across a range of tasks. Trained on over 2750 datasets, these models show significant improvements in quality, even with long input contexts, and provide better-calibrated confidence scores. The models, built on the Mixtral-8x7B and Llama3-8B bases, underwent a two-phase training process to enhance their capabilities on both short and long context tasks. They support tasks across various domains and include non-public datasets to ensure generalization to real-world settings. RefuelLLM-2 and RefuelLLM-2-small are accessible via Refuel Cloud and open-sourced on Hugging Face, with the latter available under a CC BY-NC 4.0 license. The development process involved collaboration with several open-source projects and relied on substantial computational resources provided by partners like Databricks and GCP.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 28 2,643 305 124 -22%
AI Model Fine-tuning 2 415 91 58 -44%
Reinforcement learning 2 No monthly metrics for this publish month.
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.