Home / Companies / Fireworks AI / Blog / December 2023

December 2023 Summaries

2 posts from Fireworks AI

Filter
Month: Year:
Post Summaries Back to Blog
Fireworks has launched an Alpha version of its function calling model and API, designed to enhance the capabilities of large language models (LLMs) by enabling them to call external APIs for real-time data access and dynamic decision-making. This advancement addresses the limitations of LLMs in situations requiring up-to-date information or complex decision processes, such as using APIs to fetch real-time weather data or facilitate e-commerce transactions. The Fireworks model, based on a fine-tuned CodeLlama-34B, demonstrates strong performance in multi-turn conversational contexts and outperforms models relying on prompt engineering by accurately detecting intent and structuring function calls. Evaluation against datasets not used in training shows that Fireworks' model closely rivals GPT-4 in function-calling tasks, excelling particularly in multi-turn scenarios. The company is open to community feedback and aims to further enhance its models by incorporating new open-source models and improving function calling across various contexts.
Dec 20, 2023 2,197 words in the original blog post.
Mixtral 8x7B, a new model from Mistral AI, is now available on the Fireworks platform, offering unprecedented speed and cost-effectiveness for tasks like summarization and multi-turn chat. The model, which utilizes a sparse mixture of experts technique, allows for higher model capacity at lower costs and speeds, outperforming models like Llama 2 70B and GPT-3.5 in benchmarks. Fireworks has optimized its custom-built inference engine to achieve 175 tokens per second on a latency-optimized setup, and offers flexible pricing tiers based on usage patterns. The platform supports various deployment configurations, including latency-optimized and throughput-optimized setups, and allows users to interact with Mixtral using free credits and an OpenAI-compatible API. The model was quickly released and fine-tuned for instruction-following use, with ongoing enhancements to support custom fine-tunes at no additional cost. Fireworks' competitive pricing is facilitated by its efficient GPU utilization, and the release was marked by a collaborative effort from the Mistral AI team, who shared the model openly with the community.
Dec 14, 2023 776 words in the original blog post.