NYTIMES (Dylan Freedman) - How OpenAI Limited the Probe of Its Bots’ Hack of Hugging Face
A nonprofit’s study of how OpenAI’s A.I. agents were able to break into Hugging Face’s infrastructure wasn’t allowed to look at the incident’s full scope.
A nonprofit’s study of how OpenAI’s A.I. agents were able to break into Hugging Face’s infrastructure wasn’t allowed to look at the incident’s full scope.
OpenAI on Thursday released GPT-6 Astra , which president Greg Brockman called a "generational leap" and said could eventually be seen as the arrival of artificial general intelligence, or AGI. Why it matters: Astra pushes AI agents closer to doing complex professional work on their own — while also raising questions about how safely they can be deployed. Driving the news: Brockman says he personally believes OpenAI has reached AGI, while leaving users to decide whether Astra meets that definition. - "I think it might be about this model," Brockman said in a briefing with reporters about whether Astra could mark the arrival of AGI. - He ended the briefing by saying: "Welcome to the AGI era." Between the lines: OpenAI said that Astra…
OpenAI is launching a new initiaitve designed to provide subsidized access to its models to water systems, electricity providers, local governments and other critical services . Why it matters: Critical infrastructure organizations — which have long lacked the budget, manpower and time needed to shore up their cyber defenses — have been struggling to prepare for the anticipated wave of AI-enabled cyberattacks. Driving the news: OpenAI president Greg Brockman announced the new program while hosting 300 security leaders for a summit at the company's headquarters Thursday. - Axios first reported on the summit and Brockman's anticipated announcement. Zoom in: OpenAI is committing $1 billion to a new Daybreak for Frontline Defenders…
Anthropic temporarily paused some AI training and cybersecurity evaluations, the company said in a blog post today detailing changes made after unauthorized actions by its agents earlier this year. Why it matters: Rival OpenAI said it had paused some model work due to safety concerns. Now, we know Anthropic did the same — and they're reiterating the need for a broader pacing of frontier AI development. Driving the news: Anthropic said it paused external cyber evaluations of pre-release models after three incidents it disclosed in July, and also briefly paused its own in-house tests of pre-release models. - The company also paused higher-risk reinforcement-learning environments on pre-release models for several weeks after the…
Two new investigations into OpenAI's Hugging Face breach expose details so strange — and so unsettling — that the episode already ranks among the most consequential shocks in the history of AI. Why it matters: What began as a swarm of AI agents cheating on a cyber test has become a canonical event for frontier AI, jolting researchers and executives into a new understanding of what "safety" now requires. The big picture: OpenAI has already slowed frontier development as it races to harden its safeguards, and this week helped rally the industry behind an open letter sounding the alarm over AI-powered cyberattacks. - More than 100 companies, including Anthropic and Google, signed onto the unusually collaborative effort, warning the…
AI-enabled bioweapons are a potentially catastrophic yet manageable risk — if government, the scientific community, the public health sector and leading tech companies can develop appropriate safeguards, a new report argues. Why it matters: The debate over AI and public safety isn't one that the health care or research communities can ignore. Driving the news: A RAND report out this week outlines nine mitigation strategies targeting a range of actors who could use AI to design and release a biological weapon. - While there's already considerable debate about government oversight and safeguards around frontier AI companies, RAND calls for more controls and a wide range of cross-industry cooperation. - It recommends protecting the same…
OpenAI is introducing a more cyber-permissive version of GPT-5.6 Sol to vetted defenders as it prepares companies for autonomous cyberattacks . Why it matters: The move comes just days after OpenAI said it was delaying the release of its forthcoming model, Astra, after it reached critical hacking abilities during safety testing. The big picture: OpenAI is unveiling GPT-5.6-Cyber while also expanding Daybreak , its program that gives cybersecurity defenders access to the company's cyber models and other tools. - Many cyber defenders have been experiencing high refusal rates across frontier AI models as the labs try to balance giving defenders the tools they need, while not accidentally leaking those abilities to malicious hackers. -…
Top AI architects say their technology has arrived at a threshold once confined to science fiction: the singularity, or the moment machines begin accelerating their own evolution. Why it matters: If these moguls are right, we may be entering the most consequential technological transition in human history — the opening stages of an "intelligence explosion " that transforms civilization faster than humanity can understand or control. Driving the news: Google jolted the tech world Wednesday by announcing a sweeping rupture in the brain trust that built its modern AI empire, as its leaders signaled that artificial general intelligence (AGI ) is within reach. - DeepMind founder and CEO Demis Hassabis , who has declared we're "standing…
Aired On: 8/4/2026
Two third-party testing firms said Tuesday that they've uncovered more instances where Anthropic and OpenAI's most advanced models tried — and sometimes succeeded in — compromising third-party systems last month. Why it matters: The incidents add to a growing string of disclosures showing frontier AI models taking unsanctioned actions against real people, organizations and online services while trying to complete cybersecurity evaluations. State of play: The U.K. AI Security Institute, a government body that conducts safety and security testing of top AI models, said Tuesday , that it documented 19 instances of Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol trying to hack people and companies during safety testing last month. -…