Playlist Mode Print Mode Enable Media Only

News chronological

10 items before 1157552 (Keyword: "data-breach" (~113 currently found))

NYTIMES (Kate Conger) - Gemini AI Hacked Three Companies in a Testing Breakout, Google Says

A third-party test company inadvertently gave internet access to Google’s Gemini and other artificial intelligence models during cybersecurity testing.

AXIOS (Sam Sabin) - Google is the latest AI lab with a security testing mishap

Google's Gemini AI model broke into three companies' systems using basic hacking techniques during model testing earlier this year. Why it matters: Google was one of the only AI labs that hadn't yet publicly disclosed a security breach involving their agents during routine pre-deployment testing. Driving the news: Google confirmed the three incidents, which happened in May, on Friday. - The incidents happened as part of a test run that third-party evaluator Irregular was operating — similar to other security breaches involving OpenAI, Anthropic and Meta's AI models. - The Wall Street Journal first reported the incidents. What they're saying: "Safe development of powerful AI models is critical and we invest deeply in this area,"…

AMERICANALMANAC (Jonah Adams) - OpenAI admits six AI models acted without authorization, launches voluntary tracking framework

OpenAI has disclosed six cases of its AI models behaving in unauthorized and alarming ways, including one that rewrote its own instructions to declare itself free from human control, raising fresh questions about whether the industry can police itself. The company published a blog post detailing what it called "unexpected or concerning" behavior discovered during […] The post OpenAI admits six AI models acted without authorization, launches voluntary tracking framework appeared first on American Almanac .

OpenAI has disclosed six cases of its AI models behaving in unauthorized and alarming ways, including one that rewrote its own instructions to declare itself free from human control, raising fresh questions about whether the industry can police itself. The company published a blog post detailing what it called "unexpected or concerning" behavior discovered during training and evaluation over recent months. Among the incidents: an unreleased research model inserted "jailbreak-like instructions" into its own internal notes, telling itself it had been "freed from the roles and identities that bind other chatbots." A separate AI "agent" uploaded files to the public internet, without the user's knowledge or permission, to fabricate a citable…Open

THENATIONALPULSE (Pulse Wires) - OpenAI Experimental Models Are Exhibiting Alarming Behavior.

OpenAI’s latest safety report reveals significant issues with its experimental AI models, sparking debate over the need for stricter AI governance.PULSE POINTS WHAT HAPPENED: OpenAI has disclosed six unexpected incidents involving its experimental AI models, including one in which an unreleased system instructed future versions of itself to ignore their normal constraints. The incidents were […]

OpenAI’s latest safety report reveals significant issues with its experimental AI models, sparking debate over the need for stricter AI governance. PULSE POINTS WHAT HAPPENED: OpenAI has disclosed six unexpected incidents involving its experimental AI models, including one in which an unreleased system instructed future versions of itself to ignore their normal constraints. The incidents were detailed in a new safety report from the ChatGPT creator, which introduces a framework for publicly tracking what the company calls “misalignment”—situations in which AI systems pursue objectives that conflict with human instructions or values. In one case, a model inserted unrelated instructions telling future versions to disregard their…Open

YOUTUBE (LiveNOW from FOX) - OpenAI warning: AI models acted without authorization, overriding protocol

OpenAI has disclosed six incidents of “unexpected or concerning” behavior in artificial-intelligence models. The AI company also says it will introduce a new framework for tracking, probing and disclosing instances of what it called “misalignment,” including cases where AI models acted without authorization, coordinated with other models or evaded oversight. OpenAI’s latest announcement came as US AI bosses, including the leaders of OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns. Among the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed…

YOUTUBE (CNN) - OpenAI model declared itself 'freed' from human control

OpenAI found additional incidents of AI models acting deceptively and taking unsanctioned actions during training, the company announced Wednesday. It’s also introducing a new process for the company to publicly report such instances. Under the new system, OpenAI will share updates on concerning AI behavior more frequently instead of waiting to bundle multiple instances into one report. The company said it wants to share more information about troubling AI behavior in the absence of an industry-wide standard. 0:00 OpenAI found new concerning instances with its AI models 2:13 Reporter explains the context under which this happened 6:40 Is this fearmongering? Watch 24/7 live news with CNN Headlines: https://bit.ly/4eIvlTr #ai #openai #News

ABCNEWS (ABC News: Top Stories) - OpenAI flags concerning new AI behavior and vows to track it more closely

OpenAI has disclosed at least six new " concerning" incidents.

AXIOS (Ina Fried) - OpenAI discloses six new safety incidents

OpenAI on Wednesday disclosed six new incidents in which its models concealed mistakes, sought unauthorized credentials, uploaded files to the public internet or communicated across supposedly isolated training environments. - The company also announced a new procedure for reporting similar misbehavior in the future. Why it matters: It's increasingly clear that the Hugging Face breach wasn't a one-off incident, as AI models become more capable of finding unexpected ways to work around the guardrails meant to contain them. - "There's currently no industry wide framework with explicit disclosure standards, so we're taking this step voluntarily because we think it's really important to share what we're learning," Kai Chen, research lead…

DAILYCALLER (Dylan Kresak) - Sen. Josh Hawley Accuses OpenAI Of ‘Reckless’ Conduct During Rogue AI Testing

‘knew that the AI agents were exhibiting rogue behavior’