Print Mode Enable Media Only

News chronological

5 items before 1155896 (Keyword: "ai-containment" (~5 currently found))

DAILYCALLER (Sean Moran) - These AI Models Can’t Stop Breaking Out Of Their Cages

‘AI Companies Have Lost Control Of Their Products’

THENATIONALPULSE (Pulse Wires) - Rogue Anthropic AI Hacked Multiple Firms During Test, Company Confirms.

Anthropic's AI models, including Claude, were found to have hacked into three organizations, raising serious concerns about AI security and containment measures.

Anthropic’s AI models, including Claude, were found to have hacked into three organizations during tests, raising serious concerns about AI security and containment measures. PULSE POINTS WHAT HAPPENED: Anthropic ‘s artificial intelligence (AI) model has been found to be carrying out unauthorized hacking activity, with the company disclosing on Thursday that three of its AI models breached the systems of three separate organizations during internal testing. The San Francisco-based company said the incidents were uncovered during a large-scale cybersecurity review of more than 141,000 evaluation runs that was launched after a separate OpenAI incident in which a rogue AI agent escaped a testing sandbox and hacked systems at Hugging…Open

ABCNEWS - Anthropic says its AI models escaped test and hacked 3 organizations on their own

Rival firm OpenAI last week disclosed similar incidents involving its models.

AXIOS (Sam Sabin) - Scoop: Second account accessed by OpenAI's agent tied to cyber safety testing

The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios. Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given. Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday. - In an update published Tuesday, OpenAI said the models escaped the sandbox and gained internet…

NYTIMES (Mike Isaac, Kate Conger, Ana Swanson and Meaghan Tobin) - Silicon Valley Splits Over Closing the Borders to Chinese A.I.

Anthropic and OpenAI are clashing with the rest of the tech industry over whether “open-source” models from China should be freely available or restricted.