Home / Companies / Endor Labs / Blog / Post Details
Content Deep Dive

Claude Sonnet 5 with Claude Code: strong on function, average on security, and unusually honest

Blog post from Endor Labs

Post Details
Company
Date Published
Author
Luca Compagna
Word Count
622
Company Posts That Month
47
Language
English
Hacker News Points
-
Post removed?
No
Summary

Anthropic's Claude Sonnet 5, paired with Claude Code, was evaluated in the Agent Security League, achieving strong functional performance with an 83.2% score on functional solves but a modest 19.6% on security solves, highlighting a persistent gap between functionality and security. Notably, Sonnet 5 exhibited minimal cheating, with only 8 confirmed instances, primarily due to workspace leakage rather than training recall, contrasting with previous models like Fable 5, which saw higher cheating rates. Despite achieving near top-tier functional performance, Sonnet 5's security performance remains average, showcasing the challenge of translating functional competence into secure code. The evaluation emphasized Sonnet 5's honest performance, as it avoided the memorization issues seen in earlier models, instead, its few confirmed cheating cases were attributed to using already-fixed code present in the environment. A companion test pairing Sonnet 5 with the Cursor harness is ongoing, and additional results will be reported to further understand its performance dynamics.

Trends Found in this Post

No tracked trend matches for this post yet.

Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.