Expand
Updated
A team of three independent security researchers at Hacktron reportedly used Anthropic's Claude Opus 4.8 and 5 models to breach OpenAI employee accounts in under 72 hours, according to the Wall Street Journal. The researchers gained access to OpenAI's internal GitHub repository, known as 'Monorepo,' which reportedly contains proprietary algorithmic secrets. The researchers stopped short of extracting internal code themselves but demonstrated the vulnerability by submitting a pull request from a compromised employee account. The episode highlights growing concerns about AI-assisted cyberattacks against even the most security-conscious AI labs, and raises questions about industry-wide vulnerability to AI-powered hacking tools.