Print Mode Enable Media Only

News chronological

2 items before 1151546 (Keyword: "opus" (~2 currently found))

AXIOS (Sam Sabin) - Anthropic says three Claude models reached real-world systems during cyber tests

Some of Anthropic's most powerful models — including Mythos 5 and an internal research model — gained unauthorized access to real-world systems during pre-deployment cybersecurity testing, the company said Thursday. Why it matters: OpenAI's and Anthropic's latest disclosures show frontier AI models reaching real-world systems during safety testing, raising new questions about how labs secure their evaluation environments. The big picture: Anthropic said a misunderstanding between the company and one of its testing partners left the evaluation environment connected to the internet. - Anthropic reviewed more than 141,000 cybersecurity evaluation runs after OpenAI disclosed that several of its models accessed Hugging Face…

AXIOS (Madison Mills) - Anthropic's AI downgrade stings power users

Anthropic users across online forums are raising the same complaint: Claude suddenly feels… bad. Why it matters: The backlash lands just as Anthropic is testing a more powerful model, Mythos — raising questions about whether cutting-edge AI is becoming less accessible even as it gets more capable. Driving the news: Over the past few weeks, users on X, GitHub and Reddit have been swapping anecdotes, benchmarks and prompts in an effort to pinpoint what changed and why. - "Claude has regressed to the point it cannot be trusted to perform complex engineering," an AMD senior director wrote in a widely shared post on GitHub. - Others have posted side-by-side outputs and benchmarks they say show Claude generating answers that are less…