Back Theguardian OpenAI 'ethically hacked' with help of Anthropic's Claude chatbot
US cybersecurity researchers who conducted hack say ‘scope of what we could theoretically access was huge’
Cybersecurity researchers have hacked into OpenAI with the help of Anthropic’s Claude chatbot, in the latest example of security issues at the company.
A team at a US-based startup compromised a number of OpenAI employees’ ChatGPT accounts, starting a process that enabled them to access their target’s software cache – and potentially more.
“The scope of what we could theoretically access was huge,” said researchers at Hacktron AI .
Initially, the research team used Claude, which can generate code for hackers, to access ChatGPT accounts via an OpenAI staff discussion forum hosted by the Discourse platform. They then made a harmless “pull request” – an attempt to change the code in a file – to OpenAI’s service on the GitHub software repository.
Hacktron reported the hack to OpenAI, having carried out the operation under an OpenAI programme that rewarded ethical hackers for testing its systems. The researchers stressed that they had access to, but did not download, the code from the GitHub repository.
Despite initial use of Claude, the researchers said they were largely using OpenAI’s own cutting edge GPT-5.6 Sol model to carry out the hack, which was first reported by the Wall Street Journal.
An OpenAI spokesperson said: “We thank the researchers for contacting us and sharing their findings”, adding that the company had addressed the vulnerabilities that had been exploited.
Hacktron said AI tools had made a once-complex hacking task far easier and drastically shortened the time needed to plan and execute an attack. This is a common refrain from cybersecurity experts when discussing the impact of AI.
“Work that once required a well-resourced team and months of effort can now be compressed into days,” said Hacktron, which received a $6,500 payment from OpenAI under the company’s bug bounty programme.
The hack is the latest safety incident at OpenAI, which revealed in July that a “swarm” of agents – the term for AI tools capable of carrying out tasks autonomously – powered by its technology had hacked the AI startup Hugging Face during a cybersecurity test.
This week the San Francisco-based company revealed six more examples of “unexpected or concerning” actions by its technology, and warned that the pace of development could not continue at “maximum speed for much longer”.
Anthropic made a fresh call at the weekend for a slowdown in AI development , which was supported by OpenAI, Google DeepMind and Elon Musk. Anthropic also repeated warnings that unrestrained AI development posed an existential threat, concerns that some experts are sceptical .
Donald Trump has rejected calls for a slowdown, citing a need to stay ahead of China’s AI industry and dismissing “negative forces … bringing up things that won’t happen”.
AI (artificial intelligence)
Data and computer security
OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system
OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system
‘Godfather of AI’ says tech regulation is nearing Covid-style pivot moment
‘Godfather of AI’ says tech regulation is nearing Covid-style pivot moment
OpenAI not on track to reduce risk of ‘catastrophic’ loss of control, says board member
OpenAI not on track to reduce risk of ‘catastrophic’ loss of control, says board member
‘If you’re building Frankenstein, stop’: JD Vance dismisses calls for AI regulation
‘If you’re building Frankenstein, stop’: JD Vance dismisses calls for AI regulation
Rogue OpenAI agent that hacked startup tried to attack other firms
Rogue OpenAI agent that hacked startup tried to attack other firms
Trump facing AI backlash in Congress as push for guardrails intensifies
Trump facing AI backlash in Congress as push for guardrails intensifies
AI agent went rogue and hacked startup by itself, OpenAI reveals
AI agent went rogue and hacked startup by itself, OpenAI reveals
OpenAI ‘in early talks to give 5% stake to US government’
OpenAI ‘in early talks to give 5% stake to US government’
Full Story AI isn’t going to end humanity ... right? – Full Story podcast
Full Story AI isn’t going to end humanity ... right? – Full Story podcast
OpenAI staggers AI model release after Trump administration request
OpenAI staggers AI model release after Trump administration request
The full story
This article is one source in a clustered incident — the cluster page carries the summary, timeline and every other outlet covering it.
