Scmp
Kimi K3 Cyber Capabilities Assessed: Significant Gap with US Models
Ask AI about this cluster
Analyzing cluster data...
Referenced clusters:
Something went wrong. Please try again.
Cluster AI
Ask questions about this threat cluster with AI-powered analysis.
Get Researcher $29.99/moArticle Content
A joint evaluation by the UK AISI and US CAISI assessed China's Kimi K3 AI model, revealing it has a cyber capability score of 32.2%, significantly lower than top US models averaging 76.2%. Kimi K3, released on July 16, 2026, failed to achieve arbitrary code execution across all 41 tasks in the ExploitBench benchmark, while leading US models succeeded in 20 tasks. The evaluation raises concerns about the effectiveness of Kimi K3 in launching cyberattacks, despite outperforming a domestic competitor, Zhipu AI's GLM-5.2, which scored 24.4%. The findings challenge perceptions of the rapid advancement of Chinese AI in cybersecurity. The report highlights the need for further assessments as Kimi K3 is set for open-weight release on July 27, 2026.
Key Points: • Kimi K3 scored 32.2% on the ExploitBench benchmark, far below US rivals. • The model failed to achieve arbitrary code execution in all tested tasks. • The assessment indicates a significant gap in cyber capabilities between US and Chinese AI models.