February 2024 Summaries
2 posts from Humanloop
Filter
Month:
Year:
Post Summaries
Back to Blog
Large language models (LLMs) are increasingly being used by companies to enhance product experiences and internal operations, marking a shift in the computing landscape. Evaluating LLMs presents unique challenges due to their complexity and the subjective nature of their outputs, differing from traditional software and machine learning models. The evaluation process involves various components such as prompt templates, data sources, and memory, all of which require careful configuration. Testing LLMs often focuses on integration and end-to-end tests instead of unit tests, due to factors like randomness, subjectivity, and scope. Observability and monitoring are evolving to suit the needs of LLM applications, which benefit from rapid iteration and input from diverse teams. Evaluation strategies include leveraging human, model, and heuristic judgments, with model judgments gaining prominence due to their scalability. High-quality datasets are crucial, and can be sourced from real user interactions or synthesized using LLMs. This dynamic field continues to advance, with future developments expected in AI-based evaluators, multi-modal applications, and complex agent-based workflows.
Feb 06, 2024
3,932 words in the original blog post.
Humanloop has achieved SOC 2 Type II compliance, a significant milestone that underscores the company's commitment to data security and privacy. This certification, trusted by industry leaders and essential for enterprises dealing with software providers, involves a comprehensive audit over several months to ensure effective implementation and maintenance of security controls. The audit was conducted by Insight Assurance, the same team that vetted OpenAI, with support from Vanta, highlighting the importance of SOC 2 compliance in building partnerships between enterprises and SaaS providers. Humanloop's CEO, Raza Habib, who has a background in machine learning and was featured in Forbes' 30 Under 30 list, has led the company in developing AI solutions for major technology firms. For further details on Humanloop's security policies, individuals are encouraged to visit their Trust Center or contact them directly.
Feb 05, 2024
360 words in the original blog post.