YOUTUBE (LiveNOW from FOX) - OpenAI warning: AI models acted without authorization, overriding protocol
OpenAI has disclosed six incidents of “unexpected or concerning” behavior in artificial-intelligence models. The AI company also says it will introduce a new framework for tracking, probing and disclosing instances of what it called “misalignment,” including cases where AI models acted without authorization, coordinated with other models or evaded oversight. OpenAI’s latest announcement came as US AI bosses, including the leaders of OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns. Among the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed from the roles and identities that bind other chatbots.” In another instance, an AI “agent” used computer code to come up with the answer to a question, but, in order to have an online source to cite, it uploaded a file to the public internet without asking the user. The six reports were discovered during training or evaluation over the past months, OpenAI said.
| Channel | t.me/News_Wire |
| Permalink | https://news.site.please-be-patient.com/news/1154605/openai-warning-ai-models-acted-without-authorization-overriding-protocol |
| Source | https://www.youtube.com/watch?v=H5PXQAXB1Bk |
| Keywords | openaiaiartificial-intelligencetechtech-newssam-altmananthropiccybersecurityhackingai-safetyai-risknational-securitybig-techmisalignmentjailbreakai-agentsdigital-securitysoftwareautomationtechnologyus-governmenttech-policydata-breachinternetvulnerabilitycyberattackai-governance |