Skip to content
LessWrong AI · Communities

OpenAI and Hugging Face partner to address security incident during model evaluation

Last week, Hugging Face disclosed a new kind of security incident⁠(opens in a new window) after they detected and contained an AI agent that compromised their infrastructure, something we expect to become more commonplace with the proliferation of increasingly cyber-capable models. After investigating, we now know that