Benchmarking Claude Fable 5 Reveals Harness Impact on Security Outcomes

Benchmarking Claude Fable 5 Reveals Harness Impact on Security Outcomes

First seen 18 Jun 2026, 22:54 UTC RedditEndorlabs 80% similarity 39.6

Article Content

Browse articles
ThreatCluster

Endorlabs benchmarked the Claude Fable 5 AI model using two different harnesses, revealing significant differences in security outcomes. The Cursor harness achieved a 72.6% FuncPass and 29% SecPass, while the Claude Code harness resulted in a 59.8% FuncPass and 19% SecPass. The results indicate that the agent harness has a more substantial impact on security outcomes than the model itself. Despite the improvements, the security scores remain below 30%, indicating that many vulnerabilities are still left unaddressed. This benchmarking exercise involved 200 real-world vulnerability-fixing tasks in actual projects, highlighting the importance of the agent scaffolding in achieving better security results. The findings prompt further investigation into the relationship between model capabilities and harness effectiveness.

Key Points: • Cursor harness with Claude Fable 5 achieved 72.6% FuncPass and 29% SecPass. • Claude Code harness resulted in lower scores: 59.8% FuncPass and 19% SecPass. • The agent harness significantly influences security outcomes more than the model itself.

ThreatCluster AI How this analysis works

Timeline

2017-08-18
CVE-2017-12440 published
Vulnerability assigned a CVE identifier and published in the National Vulnerability Database.
MITRE
2020-07-20
CVE-2020-15118 published
Vulnerability assigned a CVE identifier and published in the National Vulnerability Database.
MITRE
2024-04-16
CVE-2024-3571 published
Vulnerability assigned a CVE identifier and published in the National Vulnerability Database.
MITRE
2026-06-17
Reddit discussion on Claude Fable 5 performance
A Reddit post summarized the performance of Claude Fable 5 with different harnesses, emphasizing the importance of the agent harness.
Reddit
2026-06-18
Endorlabs benchmarks Claude Fable 5
Endorlabs ran tests on Claude Fable 5 using Cursor and Claude Code harnesses, revealing significant differences in security outcomes.
Endorlabs

Community

Browse all →