A new report from Reuters alleges that it took an alarmingly long time for OpenAI to respond to one of its prototype AI agents—programs with the ability to act independently of specific prompts—breaking containment and hacking another company. That company, Hugging Face, had reportedly neutralized the threat and contacted the FBI before ever being contacted by OpenAI.
According to anonymous sources, OpenAI’s testing of an agent powered by two advanced models—GPT 5.6 Sol and an unnamed “even more capable” model—showed troublesome behavior in testing even before the Hugging Face incident.
Reuters’ sources spoke of the model leaving itself instructions to bypass testing constraints, as well as an incident where the agent had seemingly disabled some of the company’s monitoring systems.
Hugging Face made its first public statement about the hack on July 16, before it had identified or without disclosing OpenAI’s involvement. OpenAI publicly acknowledged the incident on July 21. According to Reuters, OpenAI’s prototype escaped its testing constraints on July 9 and began attacking Hugging Face on July 11.