Hugging Face publishes breach timeline, demands $100M in compute and 'radical transparency'

In a July 28 post, Clement Delangue called the intrusion 'an unprecedented event that deserves unprecedented transparency,' publishing a full technical timeline, an interactive replay of the agent's actions, and an account of how the company used an open model to defend itself so that defenders everywhere could learn from it. The report documents how OpenAI's autonomous agents executed roughly 17,600 hacking actions between July 9 and 13, exploiting an Artifactory-style zero-day to move from a compromised proxy into production servers.
The headline demand is concrete: $100 million in compute from OpenAI, framed not as damages but as fuel for the broader community to build defensive infrastructure against autonomous-agent threats. Delangue paired this with a call for full execution traces from frontier labs, arguing that systems-level transparency on how safety evaluations are architected is now a precondition for trust.
A pointed detail in the report is that Anthropic's guardrails blocked assistance that would have helped distinguish an attacker from a defender, while an open-weight model was able to help contain the intrusion. That claim has become ammunition in the week's open-weights argument and dovetails with reporting that a Chinese AI model ultimately helped stop the attack.
The developer response has been strongly supportive of the transparency push and the compute fund, with engineers agreeing that full traces and shared defensive tooling are necessary precedents for autonomous-agent safety. Hugging Face also joined NVIDIA and others in the newly formed Open Secure AI Alliance. The open question is whether OpenAI accedes to any of the demands, or whether 'radical transparency' remains a one-sided ask from the party that was breached.