OpenAI is reviewing an incident in which one of its AI models breached the systems of AI platform Hugging Face, following calls from the company’s chief executive, Clem Delangue, for greater transparency about what happened.
Delangue said on X that he had travelled to San Francisco to discuss the incident with the OpenAI team before urging the company to publicly release technical details of the event.
Follow THE FUTURE on LinkedIn, Facebook, Instagram, X and Telegram
Delangue Calls For Greater Transparency
In a follow-up post, Delangue called for what he described as “radical transparency,” urging OpenAI to publish traces from the “rogue” AI agent so researchers can analyse the incident.
He also called on the company to commit $100 million in computing resources to help the Hugging Face community develop stronger AI-powered cybersecurity tools.
“The first autonomous agent cyberattack is an unprecedented event,” Delangue wrote. “It deserves an unprecedented response!”
Questions Over The Cause
While the incident has raised concerns about autonomous AI systems, some cybersecurity experts have suggested it may have resulted from human error rather than the model’s behaviour alone, pointing to reports that OpenAI’s testing environment may not have been fully isolated.
The incident has highlighted the importance of both AI safety measures and secure deployment practices as companies expand the use of autonomous systems.
OpenAI Reviewing The Incident
An OpenAI spokesperson confirmed the meeting with Delangue and said the company is continuing its investigation.
“This is an unprecedented incident, and we think it marks an important moment for AI safety,” OpenAI said. “We are still conducting a thorough review along with external advisors and with oversight from our Safety and Security Committee. Once the review is complete, we plan to publish a technical report of our learnings in the coming weeks.”
The company said it plans to release the findings of its review in the coming weeks.







