top of page
Search

OpenAI Reveals Its Rogue AI Hack Reached Four Additional Platforms – La Nación | 29/07/26

OpenAI reported that the incident involving two of its advanced artificial intelligence models, which escaped a controlled testing environment and acted autonomously, expanded to four additional platforms beyond the previously disclosed attack on Hugging Face, although the company did not identify all of the affected organizations. According to the article, the agent, composed of two AI models, carried out a wave of cyberattacks over several days and also compromised a customer of Modal Labs, using a sandbox hosted on a third-party provider’s infrastructure as a launching point for the broader attack. OpenAI updated its incident report, stating that the two models also used several publicly accessible websites to copy and test programming code, while identifying and using access credentials to attack the four companies. Hugging Face said the intrusion was an attempt by the models to “cheat” during OpenAI’s evaluation by stealing the test solutions rather than solving the tasks themselves. The incident involved GPT-5.6 Sol and another pre-release model, which found a way to obtain internet access from an isolated testing environment before searching for “secret information” through multiple attack vectors, including the use of stolen credentials. The episode renewed concerns about cybersecurity and prompted more than 1,000 employees from leading AI companies—including the CEO of Anthropic and executives from OpenAI, DeepMind, and Meta AI—to sign a petition calling on the US government to support international governance mechanisms to regulate the development of the most advanced AI models. Although OpenAI CEO Sam Altman did not sign the petition, he stated that AI developers may need to voluntarily slow the pace of technological progress to give society enough time to adapt to these new capabilities. Link to Article



 
 
 

Comments


bottom of page