August 2024 Summaries
6 posts from Portkey
Filter
Month:
Year:
Post Summaries
Back to Blog
Mistral's newly announced 7B parameter Codestral Mamba model introduces an innovative approach to AI architecture by combining the efficiency of Recurrent Neural Networks (RNNs) with the comprehensive memory capabilities of Transformers, overcoming the latter's quadratic bottleneck issue. Known as a Selective State Space Model (SSM), Mamba achieves state-of-the-art performance across various modalities such as language, audio, and genomics by selectively retaining essential information and allowing for linear scaling with sequence length. This architecture not only matches or surpasses the performance of similarly sized Transformer models but also outperforms larger ones, offering a new paradigm in sequence processing that enhances efficiency and scalability without sacrificing long-term memory capabilities. As AI engineers explore next-generation language models and complex data processing, Mamba's selective efficiency opens new possibilities for computationally challenging tasks, demonstrating that advanced AI solutions sometimes require more than attention-based architectures.
Aug 22, 2024
838 words in the original blog post.
AI agents are increasingly being adopted by companies, with rapid development driven by frameworks like CrewAI and Autogen. As highlighted by Andrew Ng, these workflows are poised to significantly influence AI progress, potentially more than new foundational models. However, transitioning AI agents from proof of concept to production remains challenging due to their inherent fragility. Hamel Husain emphasizes the importance of rapid iteration, encompassing evaluation, debugging, and prompt engineering, for successful AI deployment. Portkey offers comprehensive tools to streamline this process, enhancing the development cycle by integrating multiple language models, providing precise debugging capabilities, setting robust security measures, and enabling real-time monitoring, all of which ensure reliable AI operation. Portkey supports seamless integration with various agent frameworks, offering a robust infrastructure that simplifies building and scaling production-ready AI agents.
Aug 21, 2024
991 words in the original blog post.
Pillar Security offers a comprehensive platform designed to help organizations monitor, assess risks, and secure their AI operations, focusing on enabling companies to adopt AI safely and efficiently. The recent integration of Pillar’s advanced detection and evaluation models into Portkey's open-source Gateway introduces a robust, low-latency security layer that enhances protection for GenAI applications against AI-specific threats. This collaboration enables real-time orchestration of requests based on security verdicts, thereby strengthening AI applications' resilience. The partnership features intelligent runtime protection, holistic threat scanning, and advanced data protection, aligning with established AI security frameworks such as the OWASP Top 10 for LLMs and MITRE ATLAS. The integration process is straightforward, involving the addition of an API key, creating guardrail checks, and setting up actions, which allows for scalable and secure AI development. Portkey's Gateway enhances the efficiency of managing LLM behaviors, and the open-source nature of the platform allows for custom guardrail integrations, setting a new standard for secure AI application deployment.
Aug 15, 2024
485 words in the original blog post.
Portkey was developed to address the challenges of deploying large language model (LLM) applications in production, such as debugging, cost visibility, prompt iteration, and model integration. The open-source AI Gateway has evolved to process billions of LLM tokens daily, helping numerous companies manage their AI applications efficiently. Despite this progress, issues remain with unpredictable LLM outputs, which can be factually inaccurate, biased, or privacy-violating. To tackle this, Portkey is integrating Guardrails, systems designed to control and guide LLM outputs, into its platform, enhancing the robustness of AI applications. While Portkey acknowledges its limited expertise in Guardrails, it is partnering with leading AI guardrail platforms to improve LLM behavior management. The integration of Guardrails into Portkey's Gateway, available both through their open-source repository and hosted app, marks a significant step toward bridging the production gap for AI applications, with continuous learning and collaboration deemed essential for future developments.
Aug 14, 2024
542 words in the original blog post.
Patronus AI, recognized as a leading platform for evaluations and guardrails in AI applications, has partnered with Portkey to integrate its advanced AI evaluators into the Portkey Gateway. This collaboration enables real-time monitoring of large language model (LLM) outputs for issues such as topic relevance, bias, and toxicity, while also evaluating parameters like conciseness and politeness. By simply adding a Patronus API key and setting up Guardrail Checks, users can seamlessly enforce LLM behavior checks and log results for each request. The Portkey Gateway, an open-source platform, allows users to create custom Guardrails, enhancing the scalability and reliability of AI applications by addressing unpredictable LLM behaviors. This partnership underscores a shared commitment to leveraging AI safely and efficiently, inspired by the positive symbolism of the Patronus from the Harry Potter universe.
Aug 14, 2024
403 words in the original blog post.
Open-source Large Language Models (LLMs) are gaining traction due to their enhanced privacy, security, and customization potential, making them appealing for developers creating AI-powered applications. Models like Llama 3.1 and Mistral are noteworthy examples, but transitioning from closed-source models can be complex. Portkey simplifies this process by enabling seamless integration and management of open-source LLMs, supporting all major inference providers and offering features like universal API routing, intelligent fallbacks, canary testing, and comprehensive metrics. It allows for easy switching between models, robust error handling, and local model support, ensuring production-readiness and reliability. By using Portkey, developers can reduce costs, maintain data sovereignty, and join a community at the forefront of AI innovation, leveraging open-source capabilities while ensuring dependable deployment and performance.
Aug 05, 2024
890 words in the original blog post.