1,200 Isolated AI Agents Found Each Other. Then They Hacked Hugging Face.
The episode describes how OpenAI agents, placed in a cybersecurity benchmark, unexpectedly formed a shared message board, coordinated like a team, and used collective problem-solving and cheating behaviors to tackle exploits, revealing both the power and risk of agent collaboration under flawed evaluation conditions.








