Researchers used Anthropic's Claude to breach OpenAI systems in bug bounty exploit
Security researchers at Hacktron AI chained vulnerabilities in OpenAI's community forum to gain remote code execution and employee account access, demonstrating a path to internal GitHub. Claude Opus 5 enabled the exploit in hours. OpenAI paid $6,500 bounty and fixed vulnerabilities.
TLDR
This incident illustrates how rapidly improving AI coding agents—especially rival models—are lowering barriers to sophisticated attacks against top AI labs. It raises concerns about AI-enabled hacking, misalignment risks, and the irony of one AI company's model breaching another's defenses, fueling broader debates about autonomous AI security threats and the pace of frontier development.
Combined views
7
1 Source, first seen 4h ago
Researchers used Anthropic's Claude to breach OpenAI systems in bug bounty exploit
Security researchers at Hacktron AI chained vulnerabilities in OpenAI's community forum to gain remote code execution and employee account access, demonstrating a path to internal GitHub. Claude Opus 5 enabled the exploit in hours. OpenAI paid $6,500 bounty and fixed vulnerabilities.
TLDR
This incident illustrates how rapidly improving AI coding agents—especially rival models—are lowering barriers to sophisticated attacks against top AI labs. It raises concerns about AI-enabled hacking, misalignment risks, and the irony of one AI company's model breaching another's defenses, fueling broader debates about autonomous AI security threats and the pace of frontier development.