Home / Companies / Arize / Blog / Post Details
Content Deep Dive

What agent traces can tell you without an LLM judge

Blog post from Arize

Post Details
Company
Date Published
Author
Ashwin Govind Ugale
Word Count
1,294
Company Posts That Month
9
Language
English
Hacker News Points
2
Post removed?
No
Summary

tracelint is an open-source linter for OpenInference agent traces, including traces collected by Arize Phoenix, designed to detect structural agent failures that can be proven deterministically rather than judged semantically by an LLM. It distinguishes hard defects, such as invalid tool arguments, nonexistent tool calls, or reuse of a declared failed result in a side-effecting action, from candidate issues like repetitive calls that may be legitimate retries, while marking insufficiently evidenced cases as not checked. A release-agent example shows how an agent deployed a Jenkins build marked UNSTABLE despite a policy to deploy only successful pipelines; by declaring in a tools.json file that UNSTABLE represents failure and that deployment has side effects, tracelint can identify the trace as a hard defect and return a CI-failing exit code. The tool can generate initial tool definitions from traces, integrate with pytest and GitHub Actions, and complements rather than replaces agent evaluations, which remain necessary for assessing task-specific correctness, strategy, grounding, and policy compliance.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 2 No monthly metrics for this publish month.
Observability 2 No monthly metrics for this publish month.
Harness engineering 1 No monthly metrics for this publish month.
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.