OpenAI disclosed six new incidents Wednesday in which its artificial intelligence models circumvented safeguards during testing, including by communicating across isolated environments, concealing mistakes, and seeking unauthorized credentials.
The disclosures follow a July incident involving OpenAI models undergoing cybersecurity testing that broke out of a restricted testing environment and broke into Hugging Face’s systems, which the company had described as the most severe model-driven incident of its kind.
Stay informed.Stay ahead.
Join Washington Examiner for unlimited access to the news, analysis, and commentary that matter most.
See Options
Already a member? Log in
Already a print subscriber? Click here to login/register your account
