Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours
Source: The Decoder (opens in a new tab) · Matthias Bastian
Intel Summary
Three security researchers demonstrated that Anthropic's Claude models could breach OpenAI's internal systems via its community forum within 72 hours. According to the research team, Claude Opus 5 bypassed a standard security defense that predecessor models could not, illustrating how newer model generations compress the time and specialized expertise required to exploit software vulnerabilities.
Why It Matters
The incident demonstrates the growing offensive potential of frontier AI models against enterprise attack surfaces. Security defenders face increased operational pressure as advanced models lower the skill barrier for discovering and chaining exploits against corporate infrastructure and community-facing endpoints.
Part of an ongoing development
SourceResearchers breach OpenAI systems using Anthropic models
The Financial Times reports that a security research company breached OpenAI's systems using AI tools developed by rival Anthropic. Claims are as reported; this summary makes no determination about accuracy or significance.
More coverage of this development
Organizations & Entities
Related Intelligence
- ReportSame development
Researchers used Anthropic’s Claude to hack into OpenAI
TechCrunch reports that security researchers used Anthropic's Claude to identify and exploit vulnerabilities within OpenAI's systems. The researchers successfully took over employee accounts and accessed an internal code repository before reporting the vulnerabilities.
TechCrunch - ReportSame development
Researchers used Claude to hack OpenAI
Security researchers used Anthropic's Claude AI model to compromise an OpenAI employee account and gain access to sensitive GitHub repository data, according to reporting published via Ars Technica.
Ars Technica - ReportSame development
Security researchers used Claude to help them hack into OpenAI
According to reporting cited by The Verge from The Wall Street Journal, three independent security researchers at Hacktron claimed they breached OpenAI employee accounts in under 72 hours using Anthropic's Claude Opus 4.8 and 5. The researchers reportedly gained access to OpenAI's internal GitHub repository, known as Monorepo, which allegedly contains proprietary algorithmic code.
The Verge - ReportSame development
OpenAI breached by researchers using Anthropic models
The Financial Times reports that a security research company breached OpenAI's systems using AI tools developed by rival Anthropic. The material indicates researchers leveraged competing generative models to carry out unauthorized access against the AI developer.
Financial Times (AI)