Skip to content
Benchmarking Claude Fable 5 Reveals Harness Impact on Security Outcomes

Benchmarking Claude Fable 5 Reveals Harness Impact on Security Outcomes

First seen 18 Jun 2026, 22:54 UTC

Article Content

Browse articles
ThreatCluster AI
ThreatCluster June 19, 2026 at 22:24 UTC
  • Cursor harness with Claude Fable 5 achieved 72.6% FuncPass and 29% SecPass.
  • Claude Code harness resulted in lower scores: 59.8% FuncPass and 19% SecPass.
  • The agent harness significantly influences security outcomes more than the model itself.

Endorlabs benchmarked the Claude Fable 5 AI model using two different harnesses, revealing significant differences in security outcomes. The Cursor harness achieved a 72.6% FuncPass and 29% SecPass, while the Claude Code harness resulted in a 59.8% FuncPass and 19% SecPass. The results indicate that the agent harness has a more substantial impact on security outcomes than the model itself. Despite the improvements, the security scores remain below 30%, indicating that many vulnerabilities are still left unaddressed. This benchmarking exercise involved 200 real-world vulnerability-fixing tasks in actual projects, highlighting the importance of the agent scaffolding in achieving better security results. The findings prompt further investigation into the relationship between model capabilities and harness effectiveness.

Start a free Starter trial for enhanced analysis

Ask AI about this cluster

Updated 94d ago How this analysis works

Timeline

2017-08-18
CVE-2017-12440 published
Vulnerability assigned a CVE identifier and published in the National Vulnerability Database.
MITRE
2020-07-20
CVE-2020-15118 published
Vulnerability assigned a CVE identifier and published in the National Vulnerability Database.
MITRE
2024-04-16
CVE-2024-3571 published
Vulnerability assigned a CVE identifier and published in the National Vulnerability Database.
MITRE
2026-06-17
Reddit discussion on Claude Fable 5 performance
A Reddit post summarized the performance of Claude Fable 5 with different harnesses, emphasizing the importance of the agent harness.
Reddit
2026-06-18
Endorlabs benchmarks Claude Fable 5
Endorlabs ran tests on Claude Fable 5 using Cursor and Claude Code harnesses, revealing significant differences in security outcomes.
Endorlabs

More articles in this cluster (2)

Following this threat?

Track CVE-2017-12440 in your own feed — you're alerted when they show up in new reporting, leak sites or exploitation.

Free account · no card needed