Europe and the United Kingdom are fine-tuning their approach to AI model testing as a deadline looms for the U.S. government to set rules of the road. Why it matters: U.S. allies for years have been grappling with the same AI safety questions the Trump administration now faces. - Wherever the U.S. lands in its AI framework, architects of the EU and U.K. approaches say it would be just the start. The big picture: President Trump began his second term by scrapping his predecessor's AI strategy and pursuing an anti-regulatory regime. - But as AI models became increasingly powerful fast, the Trump administration has found itself developing an oversight framework and deciding exactly what should be covered. - Under Trump's executive…
The pace of AI development combined with soaring compute costs is squeezing the AI researchers responsible for evaluating frontier models — just as those models' capabilitie s are becoming harder to measure. Why it matters: When safety testing can't keep pace, models capable of hacking companies or aiding in the development of bioweapons could reach the public before anyone knows what they can do. - Last week's breach of Hugging Face, carried out autonomously by OpenAI's models in the middle of safety testing, shows that some of the highest-risk behaviors can emerge during pre-release testing itself. Several challenges are tying up AI safety and security researchers just as U.S. frontier AI companies race to get new models to market: …
The old ways of testing and evaluating new frontier AI models need a rewrite. Why it matters: AI models are outgrowing the existing methods of testing and benchmarking their hacking abilities — and without new tests, policymakers and corporate security teams won't have a clear way to predict what these models can actually do or whether they can be deployed safely. Driving the news: Federal agencies have until Aug. 1 to establish a classified benchmarking process to assess the capabilities of frontier AI models, although the Financial Times reports those standards may arrive as soon as this week. - When Fable 5 returned last week, Anthropic said in a blog post it was creating a standardized benchmark with Amazon, Google, Microsoft and…