Are we existentially threatened by the type of AI misalignment seen in the OpenAI Hugging Face attack?
by Alex Mallen and Girish Gupta
OpenAI models recently broke through a series of security boundaries and into Hugging Face servers in order to cheat on a cyber eval. A lot of people thought it was scary because it was a clear example of AI overreaching to do something strongly unwanted[1]. Others thought it not so...
Jul 23101