Bug Hunters Used Claude to Hack OpenAI
Source: The Information (opens in a new tab) · Rocket Drew
Intel Summary
Cybersecurity researchers at startup Hacktron AI utilized Anthropic's Claude model to breach OpenAI on July 25, gaining unauthorized access to a key software repository, according to The Information.
Why It Matters
The incident highlights how frontier AI models can be leveraged as offensive cyber tools against major AI vendors, demonstrating ongoing vulnerabilities in AI software supply chains and developer infrastructure.
Part of an ongoing development
Developing storySourceResearchers breach OpenAI systems using Anthropic models
The Financial Times reports that a security research company breached OpenAI's systems using AI tools developed by rival Anthropic. Claims are as reported; this summary makes no determination about accuracy or significance.
- Confidence
- Moderate confidence
- Corroboration
- Strongly corroborated
More coverage of this development
- Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hoursThe DecoderIndependent reporting
- Security researchers used Claude to help them hack into OpenAIThe VergeIndependent reporting
- Researchers used Anthropic’s Claude to hack into OpenAITechCrunchIndependent reporting
Organizations & Entities
Topics
Related Intelligence
- ReportSame development
Researchers used Anthropic’s Claude to hack into OpenAI
TechCrunch reports that security researchers used Anthropic's Claude to identify and exploit vulnerabilities within OpenAI's systems. The researchers successfully took over employee accounts and accessed an internal code repository before reporting the vulnerabilities.
TechCrunch - ReportSame development
Researchers used Claude to hack OpenAI
Security researchers used Anthropic's Claude AI model to compromise an OpenAI employee account and gain access to sensitive GitHub repository data, according to reporting published via Ars Technica.
Ars Technica - ReportSame development
Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours
Three security researchers demonstrated that Anthropic's Claude models could breach OpenAI's internal systems via its community forum within 72 hours. According to the research team, Claude Opus 5 bypassed a standard security defense that predecessor models could not, illustrating how newer model generations compress the time and specialized expertise required to exploit software vulnerabilities.
The Decoder - ReportSame development
Security researchers used Claude to help them hack into OpenAI
According to reporting cited by The Verge from The Wall Street Journal, three independent security researchers at Hacktron claimed they breached OpenAI employee accounts in under 72 hours using Anthropic's Claude Opus 4.8 and 5. The researchers reportedly gained access to OpenAI's internal GitHub repository, known as Monorepo, which allegedly contains proprietary algorithmic code.
The Verge