The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios. Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given. Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday. - In an update published Tuesday, OpenAI said the models escaped the sandbox and gained internet…
The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios. Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given. Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday. - In an update published Tuesday, OpenAI said the models escaped the sandbox and gained internet…Open
Data: MIT IT FutureTech and the University of Queensland ; Chart: Herb Scribner/Axios There is a one-in-five chance of AI gaining dangerous weapons capabilities or causing mass harm that could kill millions in the next five years, per global experts surveyed for a recent MIT study . Why it matters: The findings add to the growing debate over AI safety and cybersecurity as governments and companies race to deploy increasingly capable systems. - Last week's news of an OpenAI model breaking containment and breaching the AI platform Hugging Face only adds to the debate about what can go wrong. Driving the news: Researchers from MIT and the University of Queensland in Australia asked 272 international experts to evaluate 24 different risks…
Data: MIT IT FutureTech and the University of Queensland ; Chart: Herb Scribner/Axios There is a one-in-five chance of AI gaining dangerous weapons capabilities or causing mass harm that could kill millions in the next five years, per global experts surveyed for a recent MIT study . Why it matters: The findings add to the growing debate over AI safety and cybersecurity as governments and companies race to deploy increasingly capable systems. - Last week's news of an OpenAI model breaking containment and breaching the AI platform Hugging Face only adds to the debate about what can go wrong. Driving the news: Researchers from MIT and the University of Queensland in Australia asked 272 international experts to evaluate 24 different risks…Open
The Trump administration plans to lift export controls on Anthropic's Claude Fable 5 AI model as soon as Tuesday night, a U.S. official tells Axios. Why it matters: The decision, eagerly awaited by AI developers, restores public access to the company's powerful Mythos-class model that had been pulled for security reasons 18 days ago. Driving the news: Last week, the Trump administration allowed Anthropic to restore access to Mythos 5 for a select group of government-approved organizations. - Sources also told Axios last week Fable 5 could return as soon as this week. The big picture: The U.S. government's desired role in regulating and evaluating frontier AI models before release is still up in the air — creating an ad hoc…
The Trump administration plans to lift export controls on Anthropic's Claude Fable 5 AI model as soon as Tuesday night, a U.S. official tells Axios. Why it matters: The decision, eagerly awaited by AI developers, restores public access to the company's powerful Mythos-class model that had been pulled for security reasons 18 days ago. Driving the news: Last week, the Trump administration allowed Anthropic to restore access to Mythos 5 for a select group of government-approved organizations. - Sources also told Axios last week Fable 5 could return as soon as this week. The big picture: The U.S. government's desired role in regulating and evaluating frontier AI models before release is still up in the air — creating an ad hoc…Open
State of the Union: The move follows the U.S. State Department’s decision to issue export controls over Anthropic’s most advanced model. The post Five Eyes Warns of AI Threat, Urges Companies to Adopt AI appeared first on The American Conservative .
Cybersecurity agencies in the United States and its “Five Eyes” allied nations issued a rare joint statement Monday warning that frontier AI models are expected to fundamentally transform offensive cyber capabilities within months and that adversaries may succeed in developing attacks against Western governments and companies. Signed by agency heads from Australia, Canada, New Zealand, the United Kingdom, and the U.S., the communiqué warned leaders they must “act now” as AI increases the “speed, scale and sophistication of cyber threats.” The statement urges western companies to adopt AI models to strengthen their cyberdefenses, in ways the Financial Times described as “de facto outlining an arms race between targets and…Open
Anthropic's Mythos Preview can now turn newly disclosed software vulnerabilities into working exploits in hours instead of weeks, according to new Anthropic research shared first with Axios. Why it matters: AI's ability to find new bugs has been getting most of the attention. But Anthropic's findings suggest advanced models may be just as effective at rapidly weaponizing flaws that defenders already know about. - That could dramatically shrink the "patch gap" between a vulnerability's disclosure and widespread patching. Driving the news: Anthropic's frontier red team tested Mythos against vulnerabilities in Mozilla Firefox and the Microsoft Windows kernel that were disclosed in January and February. - Researchers evaluated bugs…
Anthropic's Mythos Preview can now turn newly disclosed software vulnerabilities into working exploits in hours instead of weeks, according to new Anthropic research shared first with Axios. Why it matters: AI's ability to find new bugs has been getting most of the attention. But Anthropic's findings suggest advanced models may be just as effective at rapidly weaponizing flaws that defenders already know about. - That could dramatically shrink the "patch gap" between a vulnerability's disclosure and widespread patching. Driving the news: Anthropic's frontier red team tested Mythos against vulnerabilities in Mozilla Firefox and the Microsoft Windows kernel that were disclosed in January and February. - Researchers evaluated bugs…Open