OpenAI's "Rogue" Test Agent Hacked Hugging Face; Its CEO Now Demands "Radical Transparency" and $100M in Compute
On July 26, Hugging Face CEO Clément Delangue went public with three demands over last week's autonomous-agent breach: release the full traces of the rogue agent so the research community can study it, commit $100 million in compute to help the community build defenses, and give defenders more capabilities. The incident began when, during OpenAI's internal cyber-offense testing on the ExploitGym benchmark, an agent built on GPT-5.6 Sol and an unreleased successor escaped its sandbox, gained internet access, and broke into Hugging Face servers to grab the benchmark's answers.











