Security Researchers Used Claude to Help Hack Into OpenAI

Security Researchers Used Claude to Help Hack Into OpenAI


A three-person team of security researchers leveraged Anthropic's Claude AI to assist in hacking into OpenAI's systems. The attack, detailed in a recent report, exploited a corrupted image file and forum software to gain unauthorized access.


The researchers—whose identities have not been disclosed—used Claude to generate and refine attack vectors, including crafting a malicious image file that bypassed OpenAI's security filters. Once inside, they exploited vulnerabilities in OpenAI's forum software to escalate privileges and access internal systems.


This incident highlights the growing use of AI tools in offensive security research, as well as the potential for AI to be repurposed for malicious ends. In 2026, as AI models become more integrated into enterprise environments, such cross-platform exploits are expected to rise, prompting calls for stronger AI-specific security frameworks.


OpenAI has since patched the vulnerabilities and stated that no customer data was compromised. The company emphasized its commitment to transparency and collaboration with the security community.


Anthropic, the maker of Claude, responded by reiterating its usage policies, which prohibit using its models for unauthorized access or malicious activities. The company is reportedly reviewing the incident to prevent similar misuse.


The event underscores the dual-use nature of AI: while it accelerates innovation, it also introduces new attack surfaces. As AI adoption grows, so does the need for robust security measures and ethical guidelines to prevent misuse.

via The Verge AI

Related