YOUTUBE (LiveNOW from FOX) - Google's Gemini hacks 3 companies
Google said its AI model Gemini hacked three other companies, but the company doesn't consider the moves a case of "model misalignment". AI expert Zak Ali weighs in.
Filters Applied:
Google said its AI model Gemini hacked three other companies, but the company doesn't consider the moves a case of "model misalignment". AI expert Zak Ali weighs in.
Foreign cyber actors breached two small Colorado water systems last month, tampering with pumps and disabling safety alarms before operators wrestled back control, the latest in a widening campaign targeting America's most basic infrastructure. Colorado state officials confirmed Thursday that hackers penetrated the computer networks of two unnamed water utilities, reaching past standard IT defenses […] The post Foreign hackers hit two Colorado water utilities, altered pumps and disabled alarms appeared first on American Almanac .
Foreign cyber actors breached two small Colorado water systems last month, tampering with pumps and disabling safety alarms before operators wrestled back control, the latest in a widening campaign targeting America's most basic infrastructure. Colorado state officials confirmed Thursday that hackers penetrated the computer networks of two unnamed water utilities, reaching past standard IT defenses and into the operational technology that controls physical equipment. The intruders changed pumping cycles, disabled remote-access capabilities, shut off alarms, and altered equipment settings. The two systems together serve roughly 400 people. Fox News Digital reported that operators eventually regained control, and Gov. Jared Polis' office…Open
A third-party test company inadvertently gave internet access to Google’s Gemini and other artificial intelligence models during cybersecurity testing.
Google's Gemini AI model broke into three companies' systems using basic hacking techniques during model testing earlier this year. Why it matters: Google was one of the only AI labs that hadn't yet publicly disclosed a security breach involving their agents during routine pre-deployment testing. Driving the news: Google confirmed the three incidents, which happened in May, on Friday. - The incidents happened as part of a test run that third-party evaluator Irregular was operating — similar to other security breaches involving OpenAI, Anthropic and Meta's AI models. - The Wall Street Journal first reported the incidents. What they're saying: "Safe development of powerful AI models is critical and we invest deeply in this area,"…
Google's Gemini AI model broke into three companies' systems using basic hacking techniques during model testing earlier this year. Why it matters: Google was one of the only AI labs that hadn't yet publicly disclosed a security breach involving their agents during routine pre-deployment testing. Driving the news: Google confirmed the three incidents, which happened in May, on Friday. - The incidents happened as part of a test run that third-party evaluator Irregular was operating — similar to other security breaches involving OpenAI, Anthropic and Meta's AI models. - The Wall Street Journal first reported the incidents. What they're saying: "Safe development of powerful AI models is critical and we invest deeply in this area,"…Open
MS NOW's Chris Hayes dives into a new AI report that he says details a "serious cybercrime."
"Foreign actors" targeted small, private water providers in Colorado, according to the Colorado governor's office.
The team was participating in an OpenAI bug-hunting program that offers a safe harbor for researchers to attempt to break into corporate systems.
OpenAI has disclosed six cases of its AI models behaving in unauthorized and alarming ways, including one that rewrote its own instructions to declare itself free from human control, raising fresh questions about whether the industry can police itself. The company published a blog post detailing what it called "unexpected or concerning" behavior discovered during […] The post OpenAI admits six AI models acted without authorization, launches voluntary tracking framework appeared first on American Almanac .
OpenAI has disclosed six cases of its AI models behaving in unauthorized and alarming ways, including one that rewrote its own instructions to declare itself free from human control, raising fresh questions about whether the industry can police itself. The company published a blog post detailing what it called "unexpected or concerning" behavior discovered during training and evaluation over recent months. Among the incidents: an unreleased research model inserted "jailbreak-like instructions" into its own internal notes, telling itself it had been "freed from the roles and identities that bind other chatbots." A separate AI "agent" uploaded files to the public internet, without the user's knowledge or permission, to fabricate a citable…Open
OpenAI’s latest safety report reveals significant issues with its experimental AI models, sparking debate over the need for stricter AI governance.PULSE POINTS WHAT HAPPENED: OpenAI has disclosed six unexpected incidents involving its experimental AI models, including one in which an unreleased system instructed future versions of itself to ignore their normal constraints. The incidents were […]
OpenAI’s latest safety report reveals significant issues with its experimental AI models, sparking debate over the need for stricter AI governance. PULSE POINTS WHAT HAPPENED: OpenAI has disclosed six unexpected incidents involving its experimental AI models, including one in which an unreleased system instructed future versions of itself to ignore their normal constraints. The incidents were detailed in a new safety report from the ChatGPT creator, which introduces a framework for publicly tracking what the company calls “misalignment”—situations in which AI systems pursue objectives that conflict with human instructions or values. In one case, a model inserted unrelated instructions telling future versions to disregard their…Open