Home / Companies / Braintrust / Blog / Post Details
Content Deep Dive

I ran an eval. Now what?

Blog post from Braintrust

Post Details
Company
Date Published
Author
Albert Zhang, Ornella Altunyan
Word Count
1,041
Company Posts That Month
5
Language
English
Hacker News Points
-
Post removed?
No
Summary

You built an AI application, curated test examples, picked a scoring function, and ran an evaluation. Now you need to think about that score and what steps to take next to continuously improve both your AI application and evaluation process. To determine where to focus first, review 5-10 actual examples from your evaluations, inspecting the trace for each row, input, output, and scoring outcome. Analyze these examples to identify patterns around how your application is performing and whether your scoring is accurate. You can then decide whether to refine your evaluations or make changes to your application, iterating through both with rapid development loops to continuously improve performance.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
AI Model Fine-tuning 1 897 160 75 +43%
LLM 1 3,598 465 143 -7%
Vector Search 1 4,605 291 90 +25%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.