Security researchers used Claude to help them hack into OpenAI
Source: The Verge (opens in a new tab) · Stevie Bonifield
Intel Summary
According to reporting cited by The Verge from The Wall Street Journal, three independent security researchers at Hacktron claimed they breached OpenAI employee accounts in under 72 hours using Anthropic's Claude Opus 4.8 and 5. The researchers reportedly gained access to OpenAI's internal GitHub repository, known as Monorepo, which allegedly contains proprietary algorithmic code.
Why It Matters
The reported breach demonstrates the practical effectiveness of frontier LLMs as force multipliers in offensive cyber operations against top-tier tech infrastructure. The incident underscores severe supply-chain and identity security risks for AI developers, while intensifying scrutiny around model safety guardrails designed to prevent automated cyberattacks.
Part of an ongoing development
SourceResearchers breach OpenAI systems using Anthropic models
The Financial Times reports that a security research company breached OpenAI's systems using AI tools developed by rival Anthropic. Claims are as reported; this summary makes no determination about accuracy or significance.
More coverage of this development
Organizations & Entities
Related Intelligence
- ReportSame development
Researchers used Claude to hack OpenAI
Security researchers used Anthropic's Claude AI model to compromise an OpenAI employee account and gain access to sensitive GitHub repository data, according to reporting published via Ars Technica.
Ars Technica - ReportSame development
Researchers used Anthropic’s Claude to hack into OpenAI
TechCrunch reports that security researchers used Anthropic's Claude to identify and exploit vulnerabilities within OpenAI's systems. The researchers successfully took over employee accounts and accessed an internal code repository before reporting the vulnerabilities.
TechCrunch - ReportSame development
Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours
Three security researchers demonstrated that Anthropic's Claude models could breach OpenAI's internal systems via its community forum within 72 hours. According to the research team, Claude Opus 5 bypassed a standard security defense that predecessor models could not, illustrating how newer model generations compress the time and specialized expertise required to exploit software vulnerabilities.
The Decoder - ReportSame development
OpenAI breached by researchers using Anthropic models
The Financial Times reports that a security research company breached OpenAI's systems using AI tools developed by rival Anthropic. The material indicates researchers leveraged competing generative models to carry out unauthorized access against the AI developer.
Financial Times (AI)