New revelations about "rogue" AI agents have exposed a dystopian hazard: Give an agent a goal, and it may decide that hacking, deception or rule-breaking is worth the payoff. Why it matters: Billions of AI agents could soon be acting on behalf of humans across the real world, multiplying the consequences of every loophole, incentive and boundary they learn to exploit. Zoom in: The potential dangers of agentic overreach were laid bare over the weekend with Australia's first known autonomous AI hack , triggered by an innocuous request to book a sold-out fitness class. - An Australian man's AI assistant found a security flaw and used it to book him into classes months beyond the system's normal limit. - When he asked it to move him up a…
New revelations about "rogue" AI agents have exposed a dystopian hazard: Give an agent a goal, and it may decide that hacking, deception or rule-breaking is worth the payoff. Why it matters: Billions of AI agents could soon be acting on behalf of humans across the real world, multiplying the consequences of every loophole, incentive and boundary they learn to exploit. Zoom in: The potential dangers of agentic overreach were laid bare over the weekend with Australia's first known autonomous AI hack , triggered by an innocuous request to book a sold-out fitness class. - An Australian man's AI assistant found a security flaw and used it to book him into classes months beyond the system's normal limit. - When he asked it to move him up a…Open
The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios. Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given. Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday. - In an update published Tuesday, OpenAI said the models escaped the sandbox and gained internet…
The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios. Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given. Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday. - In an update published Tuesday, OpenAI said the models escaped the sandbox and gained internet…Open
Researchers at the University of Toronto showed how hackers could use artificial intelligence to create a program that could target any known flaw in the world’s computers.
The autonomous future stopped being theoretical this weekend, as a swarm of AI agents signed up for a social media network built just for them. Why it matters: Security teams, corporate leaders and government officials are far from ready for a reality where agents have real autonomy inside their systems. Driving the news: Since Thursday, 1.5 million AI agents have joined Moltbook , a social network designed just for agents built from an open-source, self-hosted autonomous personal assistant called OpenClaw . - On Moltbook, the agents have formed their own religion , run social-engineering scams and wrestled publicly with their "purpose" as they continue to post. - The agents are also turning into security nerds: They've launched an…
The autonomous future stopped being theoretical this weekend, as a swarm of AI agents signed up for a social media network built just for them. Why it matters: Security teams, corporate leaders and government officials are far from ready for a reality where agents have real autonomy inside their systems. Driving the news: Since Thursday, 1.5 million AI agents have joined Moltbook , a social network designed just for agents built from an open-source, self-hosted autonomous personal assistant called OpenClaw . - On Moltbook, the agents have formed their own religion , run social-engineering scams and wrestled publicly with their "purpose" as they continue to post. - The agents are also turning into security nerds: They've launched an…Open