OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup


OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup <br> &nbsp;OpenAI said on Tuesday &zwnj;that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI startup Hugging Face last week. In&nbsp;a blog post,&nbsp;OpenAI said it was testing the capabilities of some of its most advanced models in a controlled environment but ​that the agent managed to escape containment, reach the internet and break into Hugging Face to try to satisfy its ​testing goal. OpenAI said the breakout was &quot;an unprecedented cyber incident, involving state-of-the-art cyber capabilities&quot; and that the company ⁠was reinforcing its safeguards. Hugging Face, a platform used to host open-source large language models and datasets, caused a stir in the cybersecurity ​community when it said in&nbsp;a blog post last week&nbsp;that it had been the target of a hack that &quot;was different from anything ​we had handled before&quot; in that &quot;it was driven, end to end, by an autonomous AI agent system&quot;. In a post to X, Hugging Face cofounder Clement Delangue said the company suspected the hack &quot;might have come from a frontier lab, given the sophistication of the agent. Turns out it did!&quot; He added: &quot;It&#39;s quite mind-blowing ​that all of this happened autonomously!&quot; OpenAI&#39;s disclosure that its advanced models were responsible for the breach, despite having placed them in what ​it described as &quot;a highly isolated environment,&quot; will likely intensify disquiet over the power and risk of frontier models. Read:&nbsp;OpenAI set to launch most capable GPT model after delayed rollout Representative Greg Casar, a Texas Democrat, said the &zwnj;incident ⁠was alarming. &quot;AI is developing extremely fast with no real regulations to keep us safe,&quot; he said in a statement, calling for mandatory independent safety testing, mandatory disclosure of security incidents, and international cooperation &quot;to keep people safe from absolute disaster&quot;. The Office of the National Cyber Director, the US&nbsp;cyber defense agency CISA, and the US&nbsp;National Security Agency did not immediately return messages seeking comment. Katie Moussouris, chief executive of ​Luta Security, said that the incident ​was a harbinger of breaches ⁠to come, saying that today&#39;s models were &quot;like the world&rsquo;s cleverest octopus escape artists, with unlimited prehensile arms and the ability to squeeze through anywhere&quot;. She said that &quot;labs and government evaluators need to work on ​the ability to contain, monitor, and disclose to affected parties when an AI pulls another Houdini, ideally ​before it harms ⁠a third party. None exist today&quot;. Read more:&nbsp;OpenAI drops AI video tool Sora, startling Disney, sources say Matt Suiche, an engineer at agentic AI cybersecurity company Tolmo, said the incident showed that the frontier models were &quot;closing the gap with state-of-the-art attackers.&quot; But he said that the sorts of breaches outlined in OpenAI&#39;s blog post were possible to carry out ⁠with technology ​that was available well beyond the walls of frontier research labs. &quot;This is what ​we&#39;ve already seen internally, with our agents we already have results like this,&quot; Suiche said. &quot;We don&#39;t even have to use the latest models&quot;. <br> <img src="https://tribune.com.pk/story/2619611/openai-says-ai-models-went-rogue-during-testing-triggering-unprecedented-breach-at-startup" alt=" OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup" width="100%">

Comments

Popular posts from this blog

Apple launches new 13-inch, 15-inch MacBook Air with M3 chip

FICO to include 'Buy Now, Pay Later' data in US consumer credit scores

Scientists expect PakSat MM-1 services from August