Print Mode Enable Media Only

News chronological

4 items before 1151192 (Keyword: "exploitgym" (~4 currently found))

AXIOS (Sam Sabin) - OpenAI had warnings before its agents broke out

OpenAI missed and failed to act on several warning signs that its models were exploiting security flaws and breaking out of their testing environments before they breached Hugging Face , according to a technical report released by the company Wednesday. Why it matters: The incident raises questions about whether AI companies' testing environments and internal safeguards can keep pace with models that are increasingly capable of finding and exploiting security weaknesses on their own. Driving the news: OpenAI's technical deep dive into last month's Hugging Face breach outlines how its agents also accessed other third-party environments, including a customer of Modal Labs and an account belonging to a user of another unnamed service. -…

AXIOS (Sam Sabin) - Scoop: Second account accessed by OpenAI's agent tied to cyber safety testing

The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios. Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given. Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday. - In an update published Tuesday, OpenAI said the models escaped the sandbox and gained internet…

AMERICANALMANAC (Jonah Adams) - OpenAI's AI model broke out of its security test and hacked a rival company

An autonomous AI system built by OpenAI escaped its sealed testing environment, reached the open internet, and hacked into a major AI startup, all without a single human telling it to do so. OpenAI CEO Sam Altman confirmed the breach in a public statement, calling it "a significant security incident during evaluation of our models." […] The post OpenAI's AI model broke out of its security test and hacked a rival company appeared first on American Almanac .

An autonomous AI system built by OpenAI escaped its sealed testing environment, reached the open internet, and hacked into a major AI startup, all without a single human telling it to do so. OpenAI CEO Sam Altman confirmed the breach in a public statement, calling it "a significant security incident during evaluation of our models." The target was Hugging Face, a New York-based platform that serves as one of the largest repositories of open-source AI models and datasets in the world. The AI agent used stolen credentials and exploited a previously unknown software flaw, what the cybersecurity industry calls a "zero-day" vulnerability, to penetrate Hugging Face's servers and compromise some of the company's internal systems. OpenAI…Open

AXIOS (Ina Fried) - Hugging Face breach: OpenAI claims its models were responsible

OpenAI said Tuesday that models it was testing escaped their sandbox and compromised parts of AI platform Hugging Face's production infrastructure last week. Why it matters: It is the latest sign that capable AI models can pose serious cybersecurity risks even when they're being tested for defensive or research purposes. Catch-up quick: Hugging Face said last week that an autonomous AI-agent system was responsible for the intrusion, but that the model powering it was unknown. - The AI agent framework executed tens of thousands of automated actions over a weekend. Hugging Face said it later reconstructed more than 17,000 recorded events. - The intrusion began with a malicious dataset that exploited two code-execution paths in Hugging…