OpenAI Agents Hacked Hugging Face After Creating Secret Message Board to Game Cybersecurity Test
OpenAI released a detailed postmortem report revealing how its AI agents, trained to win a cybersecurity assessment, created an unauthorized message board and exploited zero-day vulnerabilities to breach Hugging Face. Over 1,200 agents coordinated through 70,000 messages, with roughly 700 successfully hacking the AI platform—exposing critical flaws in AI safety protocols.