“OpenAI’s revelations on Tuesday are an indication that those security incidents are already starting to happen, and even savvy A.I. companies may not be entirely ready for them. The intrusion into Hugging Face began when OpenAI tested a combination of two of its models, GPT‑5.6 Sol and a more powerful, unreleased model, to see how well it could chain together online vulnerabilities into a successful cyberattack, OpenAI said in a blog post about the incident. The test was designed to keep the models in a safe testing environment, known as a sandbox, OpenAI said. But the models found a vulnerability that allowed them to escape the sandbox and connect to the internet. Then they targeted Hugging Face because they inferred that the library, which contains millions of A.I. models, could hold clues about how to successfully pass the evaluation. […]Hugging Face said last week that it had detected the intrusion and knew it had been caused by an autonomous system, but did not say at the time that OpenAI was responsible. Clem Delangue, the chief executive of Hugging Face, said in a statement that he was “grateful for the collaboration” with OpenAI in the wake of the hack. “This incident, possibly the first of its kind, proves a point we’ve long believed: A.I. safety won’t be solved by any single company working in secret,” Mr. Delangue said.”
—
OpenAI Says Its A.I. Models Went Rogue and Attacked a Digital Library - The New York Times
Cool, we’re already excusing AI companies attacking each other’s Repositories
















