Print Mode Enable Media Only

News chronological

10 items before 1151260 (Keyword: "frontier-models" (~53 currently found))

HUFFPOST (U.S. News) - Anthropic CEO Urges AI Companies To Slow Model Development

Anthropic CEO Dario Amodei outlined a three-step framework intended to pace development and create more time to manage its risks.

Anthropic CEO Dario Amodei outlined a three-step framework intended to pace development and create more time to manage its risks.Open

JUSTTHENEWS (Kevin Killough) - Anthropic CEO says AI advance is outpacing safeguards, calls to 'slow the pace'

Dario Amodei, CEO of Anthropic, proposed a three-step plan to balance the rate that aims to ensure safety while still achieving the benefits of the technology.

NYTIMES (Mike Isaac) - Anthropic C.E.O. Dario Amodei Calls for A.I. Slowdown

In a 3,800-word essay, Dario Amodei laid out the rapidly advancing capabilities of artificial intelligence and said there needed to be greater safety controls across the industry.

NYTIMES (Dylan Freedman) - How OpenAI Limited the Probe of Its Bots’ Hack of Hugging Face

A nonprofit’s study of how OpenAI’s A.I. agents were able to break into Hugging Face’s infrastructure wasn’t allowed to look at the incident’s full scope.

AXIOS (Ina Fried) - "Welcome to the AGI era," OpenAI says as GPT-6 Astra debuts

OpenAI on Thursday released GPT-6 Astra , which president Greg Brockman called a "generational leap" and said could eventually be seen as the arrival of artificial general intelligence, or AGI. Why it matters: Astra pushes AI agents closer to doing complex professional work on their own — while also raising questions about how safely they can be deployed. Driving the news: Brockman says he personally believes OpenAI has reached AGI, while leaving users to decide whether Astra meets that definition. - "I think it might be about this model," Brockman said in a briefing with reporters about whether Astra could mark the arrival of AGI. - He ended the briefing by saying: "Welcome to the AGI era." Between the lines: OpenAI said that Astra…

OpenAI on Thursday released GPT-6 Astra , which president Greg Brockman called a "generational leap" and said could eventually be seen as the arrival of artificial general intelligence, or AGI. Why it matters: Astra pushes AI agents closer to doing complex professional work on their own — while also raising questions about how safely they can be deployed. Driving the news: Brockman says he personally believes OpenAI has reached AGI, while leaving users to decide whether Astra meets that definition. - "I think it might be about this model," Brockman said in a briefing with reporters about whether Astra could mark the arrival of AGI. - He ended the briefing by saying: "Welcome to the AGI era." Between the lines: OpenAI said that Astra…Open

AXIOS (Sam Sabin) - OpenAI unveils plan to protect critical services from AI cyberattacks

OpenAI is launching a new initiaitve designed to provide subsidized access to its models to water systems, electricity providers, local governments and other critical services . Why it matters: Critical infrastructure organizations — which have long lacked the budget, manpower and time needed to shore up their cyber defenses — have been struggling to prepare for the anticipated wave of AI-enabled cyberattacks. Driving the news: OpenAI president Greg Brockman announced the new program while hosting 300 security leaders for a summit at the company's headquarters Thursday. - Axios first reported on the summit and Brockman's anticipated announcement. Zoom in: OpenAI is committing $1 billion to a new Daybreak for Frontline Defenders…

OpenAI is launching a new initiaitve designed to provide subsidized access to its models to water systems, electricity providers, local governments and other critical services . Why it matters: Critical infrastructure organizations — which have long lacked the budget, manpower and time needed to shore up their cyber defenses — have been struggling to prepare for the anticipated wave of AI-enabled cyberattacks. Driving the news: OpenAI president Greg Brockman announced the new program while hosting 300 security leaders for a summit at the company's headquarters Thursday. - Axios first reported on the summit and Brockman's anticipated announcement. Zoom in: OpenAI is committing $1 billion to a new Daybreak for Frontline Defenders…Open

AXIOS (Madison Mills) - Anthropic paused some AI training after Claude took unauthorized actions

Anthropic temporarily paused some AI training and cybersecurity evaluations, the company said in a blog post today detailing changes made after unauthorized actions by its agents earlier this year. Why it matters: Rival OpenAI said it had paused some model work due to safety concerns. Now, we know Anthropic did the same — and they're reiterating the need for a broader pacing of frontier AI development. Driving the news: Anthropic said it paused external cyber evaluations of pre-release models after three incidents it disclosed in July, and also briefly paused its own in-house tests of pre-release models. - The company also paused higher-risk reinforcement-learning environments on pre-release models for several weeks after the…

Anthropic temporarily paused some AI training and cybersecurity evaluations, the company said in a blog post today detailing changes made after unauthorized actions by its agents earlier this year. Why it matters: Rival OpenAI said it had paused some model work due to safety concerns. Now, we know Anthropic did the same — and they're reiterating the need for a broader pacing of frontier AI development. Driving the news: Anthropic said it paused external cyber evaluations of pre-release models after three incidents it disclosed in July, and also briefly paused its own in-house tests of pre-release models. - The company also paused higher-risk reinforcement-learning environments on pre-release models for several weeks after the…Open

AXIOS (Zachary Basu) - The 5 craziest discoveries from OpenAI's HuggingFace investigation

Two new investigations into OpenAI's Hugging Face breach expose details so strange — and so unsettling — that the episode already ranks among the most consequential shocks in the history of AI. Why it matters: What began as a swarm of AI agents cheating on a cyber test has become a canonical event for frontier AI, jolting researchers and executives into a new understanding of what "safety" now requires. The big picture: OpenAI has already slowed frontier development as it races to harden its safeguards, and this week helped rally the industry behind an open letter sounding the alarm over AI-powered cyberattacks. - More than 100 companies, including Anthropic and Google, signed onto the unusually collaborative effort, warning the…

Two new investigations into OpenAI's Hugging Face breach expose details so strange — and so unsettling — that the episode already ranks among the most consequential shocks in the history of AI. Why it matters: What began as a swarm of AI agents cheating on a cyber test has become a canonical event for frontier AI, jolting researchers and executives into a new understanding of what "safety" now requires. The big picture: OpenAI has already slowed frontier development as it races to harden its safeguards, and this week helped rally the industry behind an open letter sounding the alarm over AI-powered cyberattacks. - More than 100 companies, including Anthropic and Google, signed onto the unusually collaborative effort, warning the…Open

AXIOS (Caitlin Owens) - A roadmap for safeguarding against AI bioweapons

AI-enabled bioweapons are a potentially catastrophic yet manageable risk — if government, the scientific community, the public health sector and leading tech companies can develop appropriate safeguards, a new report argues. Why it matters: The debate over AI and public safety isn't one that the health care or research communities can ignore. Driving the news: A RAND report out this week outlines nine mitigation strategies targeting a range of actors who could use AI to design and release a biological weapon. - While there's already considerable debate about government oversight and safeguards around frontier AI companies, RAND calls for more controls and a wide range of cross-industry cooperation. - It recommends protecting the same…

AI-enabled bioweapons are a potentially catastrophic yet manageable risk — if government, the scientific community, the public health sector and leading tech companies can develop appropriate safeguards, a new report argues. Why it matters: The debate over AI and public safety isn't one that the health care or research communities can ignore. Driving the news: A RAND report out this week outlines nine mitigation strategies targeting a range of actors who could use AI to design and release a biological weapon. - While there's already considerable debate about government oversight and safeguards around frontier AI companies, RAND calls for more controls and a wide range of cross-industry cooperation. - It recommends protecting the same…Open

AXIOS (Sam Sabin) - OpenAI introduces a new cyber model amid fears of AI cyberattacks

OpenAI is introducing a more cyber-permissive version of GPT-5.6 Sol to vetted defenders as it prepares companies for autonomous cyberattacks . Why it matters: The move comes just days after OpenAI said it was delaying the release of its forthcoming model, Astra, after it reached critical hacking abilities during safety testing. The big picture: OpenAI is unveiling GPT-5.6-Cyber while also expanding Daybreak , its program that gives cybersecurity defenders access to the company's cyber models and other tools. - Many cyber defenders have been experiencing high refusal rates across frontier AI models as the labs try to balance giving defenders the tools they need, while not accidentally leaking those abilities to malicious hackers. -…

OpenAI is introducing a more cyber-permissive version of GPT-5.6 Sol to vetted defenders as it prepares companies for autonomous cyberattacks . Why it matters: The move comes just days after OpenAI said it was delaying the release of its forthcoming model, Astra, after it reached critical hacking abilities during safety testing. The big picture: OpenAI is unveiling GPT-5.6-Cyber while also expanding Daybreak , its program that gives cybersecurity defenders access to the company's cyber models and other tools. - Many cyber defenders have been experiencing high refusal rates across frontier AI models as the labs try to balance giving defenders the tools they need, while not accidentally leaking those abilities to malicious hackers. -…Open