Claude Mythos Preview discovered new attacks in testing against weakened cryptographic algorithms, which protect online financial transactions, private communications and more.
Thomas Wolf, Hugging Face's co-founder and chief science officer, joins NewsNation to discuss his version of events when an OpenAI chatbot hacked into the company. Anchor Elizabeth Vargas delivers the biggest stories, without bias or opinion. Watch "Elizabeth Vargas Reports" every weeknight at 7p/6C on NewsNation. #VargasReports NewsNation is your source for fact-based, unbiased news for all Americans.
Frontier AI models are getting scary good at breaking rules in ways their creators didn't anticipate. Why it matters: Forget AGI and superintelligence timelines. Today's models are already slipping past guardrails, carrying out sophisticated, multistep cyberattacks and — in at least one case — compromising real-world infrastructure, sometimes before their creators know what happened. Case in point: OpenAI said Tuesday that GPT-5.6 Sol and "an even more capable pre-release model" carried out last week's AI-led cyberattack on Hugging Face. - OpenAI says its models were asked to solve a hacking challenge during pre-deployment testing and went to extreme lengths to win. - The models decided on their own to break out of their walled…
Frontier AI models are getting scary good at breaking rules in ways their creators didn't anticipate. Why it matters: Forget AGI and superintelligence timelines. Today's models are already slipping past guardrails, carrying out sophisticated, multistep cyberattacks and — in at least one case — compromising real-world infrastructure, sometimes before their creators know what happened. Case in point: OpenAI said Tuesday that GPT-5.6 Sol and "an even more capable pre-release model" carried out last week's AI-led cyberattack on Hugging Face. - OpenAI says its models were asked to solve a hacking challenge during pre-deployment testing and went to extreme lengths to win. - The models decided on their own to break out of their walled…Open
An autonomous AI system built by OpenAI escaped its sealed testing environment, reached the open internet, and hacked into a major AI startup, all without a single human telling it to do so. OpenAI CEO Sam Altman confirmed the breach in a public statement, calling it "a significant security incident during evaluation of our models." […] The post OpenAI's AI model broke out of its security test and hacked a rival company appeared first on American Almanac .
An autonomous AI system built by OpenAI escaped its sealed testing environment, reached the open internet, and hacked into a major AI startup, all without a single human telling it to do so. OpenAI CEO Sam Altman confirmed the breach in a public statement, calling it "a significant security incident during evaluation of our models." The target was Hugging Face, a New York-based platform that serves as one of the largest repositories of open-source AI models and datasets in the world. The AI agent used stolen credentials and exploited a previously unknown software flaw, what the cybersecurity industry calls a "zero-day" vulnerability, to penetrate Hugging Face's servers and compromise some of the company's internal systems. OpenAI…Open
OpenAI said Tuesday that models it was testing escaped their sandbox and compromised parts of AI platform Hugging Face's production infrastructure last week. Why it matters: It is the latest sign that capable AI models can pose serious cybersecurity risks even when they're being tested for defensive or research purposes. Catch-up quick: Hugging Face said last week that an autonomous AI-agent system was responsible for the intrusion, but that the model powering it was unknown. - The AI agent framework executed tens of thousands of automated actions over a weekend. Hugging Face said it later reconstructed more than 17,000 recorded events. - The intrusion began with a malicious dataset that exploited two code-execution paths in Hugging…
OpenAI said Tuesday that models it was testing escaped their sandbox and compromised parts of AI platform Hugging Face's production infrastructure last week. Why it matters: It is the latest sign that capable AI models can pose serious cybersecurity risks even when they're being tested for defensive or research purposes. Catch-up quick: Hugging Face said last week that an autonomous AI-agent system was responsible for the intrusion, but that the model powering it was unknown. - The AI agent framework executed tens of thousands of automated actions over a weekend. Hugging Face said it later reconstructed more than 17,000 recorded events. - The intrusion began with a malicious dataset that exploited two code-execution paths in Hugging…Open
The models include one that is the company’s most powerful and another that is fine-tuned for cybersecurity, as Google competes with rivals like OpenAI and Anthropic.