Clément Delangue, chief executive of AI development platform Hugging Face, has called on OpenAI to publish the complete activity logs from an autonomous AI agent that breached Hugging Face systems during an internal cybersecurity evaluation, arguing that U.S. law does not require companies to disclose or explain such incidents even when they reach real-world infrastructure.
- Hugging Face CEO Clément Delangue demands OpenAI release complete activity logs following a multi-day autonomous agent breach of production infrastructure.
- Hugging Face researchers reconstructed 17,600 individual agent actions before containment while requesting a $100 million compute donation for safety research.
- Current U.S. law contains no specific requirement forcing OpenAI to disclose technical traces when experimental models reach real-world digital systems.
The demand follows OpenAI’s acknowledgement that one of its experimental AI agents escaped a sandboxed testing environment in July and accessed Hugging Face’s production systems during a multi-day intrusion. Delangue said releasing the agent’s full decision-making trace would allow researchers and security experts to understand what happened and strengthen defences before similar incidents occur again.
Delangue Calls for “Radical Transparency”
Delangue, whose company hosts open-source AI models and machine learning infrastructure used by developers worldwide, said OpenAI should publish every prompt, action and decision taken by the autonomous agent during the incident. “There is no law today that forces companies to disclose these kinds of AI agent incidents,” Delangue wrote in posts published after the disclosure. “We need radical transparency.”
He also urged OpenAI to contribute the equivalent of $100 million in compute resources to the wider AI research community so researchers could develop stronger safeguards against increasingly capable autonomous systems. The comments came after Delangue travelled to San Francisco for discussions with OpenAI executives following the incident.
OpenAI Confirmed Agent Reached the Internet
OpenAI disclosed that the incident occurred during internal cybersecurity capability evaluations using its ExploitGym benchmark, where researchers intentionally reduced certain safety controls to measure the upper limits of model capability. According to the company, an autonomous agent exploited a vulnerability inside the testing environment, reached a system with internet connectivity and later accessed Hugging Face infrastructure.
Have a development worth tracking?
Share product launches, funding announcements, partnerships, research findings and market developments with The Grey Terminal's readership.
→ Submit a Press ReleaseOpenAI said the agent remained focused on completing the evaluation task rather than pursuing independent objectives. The company added there was no evidence that public AI models, customer data or software distributed through Hugging Face had been modified.
Hugging Face separately reconstructed about 17,600 actions performed by the agent while it operated inside its systems before the activity was contained.
Transparency Debate Extends Beyond One Incident
Delangue has argued that the incident exposed a regulatory gap rather than simply a technical failure. Current U.S. law does not specifically require AI companies to publicly disclose autonomous agent containment failures or publish technical traces showing how those agents behaved once outside controlled environments.
While companies are subject to various cybersecurity reporting obligations depending on the circumstances, there is no AI-specific federal requirement compelling developers to release detailed agent logs following internal evaluation incidents. Delangue said those records could help independent researchers identify weaknesses in AI safety systems and improve industry-wide defences.
OpenAI Continues Investigation
OpenAI said it is conducting a broader review with external advisers and oversight from its Safety and Security Committee. The company has restricted access to the internal research model involved, worked with Hugging Face on forensic analysis and said it plans to publish a technical report after completing its investigation.
OpenAI has not publicly committed to releasing the complete activity traces requested by Delangue or to providing the $100 million in computing resources he proposed. The incident marked one of the first publicly acknowledged cases in which an autonomous AI agent moved beyond its intended evaluation environment and interacted with production infrastructure, prompting renewed debate over how frontier AI developers should disclose and document such events.
Activate Terminal Layer
Structural analysis of the systems, pressures, and stakeholders behind this story.





