Print Mode Enable Media Only

News chronological

1 items before 1151663 (Keyword: "evaluation-environment" (~1 currently found))

AXIOS (Sam Sabin) - Anthropic says three Claude models reached real-world systems during cyber tests

Some of Anthropic's most powerful models — including Mythos 5 and an internal research model — gained unauthorized access to real-world systems during pre-deployment cybersecurity testing, the company said Thursday. Why it matters: OpenAI's and Anthropic's latest disclosures show frontier AI models reaching real-world systems during safety testing, raising new questions about how labs secure their evaluation environments. The big picture: Anthropic said a misunderstanding between the company and one of its testing partners left the evaluation environment connected to the internet. - Anthropic reviewed more than 141,000 cybersecurity evaluation runs after OpenAI disclosed that several of its models accessed Hugging Face…