The pace of AI development combined with soaring compute costs is squeezing the AI researchers responsible for evaluating frontier models — just as those models' capabilitie s are becoming harder to measure. Why it matters: When safety testing can't keep pace, models capable of hacking companies or aiding in the development of bioweapons could reach the public before anyone knows what they can do. - Last week's breach of Hugging Face, carried out autonomously by OpenAI's models in the middle of safety testing, shows that some of the highest-risk behaviors can emerge during pre-release testing itself. Several challenges are tying up AI safety and security researchers just as U.S. frontier AI companies race to get new models to market: …
State of the Union: Claude Fable 5 will be available for public use Wednesday. The post U.S. Lifts Export Controls on Anthropic AI Models appeared first on The American Conservative .
On Tuesday, AI company Anthropic announced that the U.S. government had lifted export controls on its newest models, following a more than two-week suspension prompted by cybersecurity concerns. Anthropic said that Claude Fable 5 would be available to consumers worldwide starting on Wednesday, although use of the more capable Claude Mythos 5 will remain limited to certain approved organizations. Access to both models had been blocked since June 12, when the Trump Administration imposed restrictions on their use by foreign nationals. The White House was concerned that Anthropic’s safeguards could be bypassed to help identify software vulnerabilities, allowing for disruptive cyberattacks at potentially large scale. In response,…Open
OpenAI rolled out a cybersecurity model that rivals the capabilities of Mythos — without nearly as much fanfare or political pushback as Anthropic received. Why it matters: The seemingly straightforward model release raises questions about what actually triggered the Trump administration's concerns about Anthropic's Fable 5 and Mythos 5. Catch up quick: OpenAI debuted an update to GPT-5.5-Cyber on Monday as part of a slew of announcements aimed at deepening its work with cybersecurity companies and researchers. Yes, but: The new GPT-5.5-Cyber achieved an 85.6% score in CyberGym, an internal benchmark that measures whether an AI agent can reproduce known software vulnerabilities. - In comparison, Mythos 5 scored 83.8% on the same…
Anthropic's Mythos Preview can now turn newly disclosed software vulnerabilities into working exploits in hours instead of weeks, according to new Anthropic research shared first with Axios. Why it matters: AI's ability to find new bugs has been getting most of the attention. But Anthropic's findings suggest advanced models may be just as effective at rapidly weaponizing flaws that defenders already know about. - That could dramatically shrink the "patch gap" between a vulnerability's disclosure and widespread patching. Driving the news: Anthropic's frontier red team tested Mythos against vulnerabilities in Mozilla Firefox and the Microsoft Windows kernel that were disclosed in January and February. - Researchers evaluated bugs…