Skip to main content
SecurityModels

Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours

Source: The Decoder (opens in a new tab) · Matthias Bastian

Intel Summary

Three security researchers demonstrated that Anthropic's Claude models could breach OpenAI's internal systems via its community forum within 72 hours. According to the research team, Claude Opus 5 bypassed a standard security defense that predecessor models could not, illustrating how newer model generations compress the time and specialized expertise required to exploit software vulnerabilities.

Why It Matters

The incident demonstrates the growing offensive potential of frontier AI models against enterprise attack surfaces. Security defenders face increased operational pressure as advanced models lower the skill barrier for discovering and chaining exploits against corporate infrastructure and community-facing endpoints.

Part of an ongoing development

Source

Researchers breach OpenAI systems using Anthropic models

The Financial Times reports that a security research company breached OpenAI's systems using AI tools developed by rival Anthropic. Claims are as reported; this summary makes no determination about accuracy or significance.

More coverage of this development

Organizations & Entities