Tensorwire
Products & tools · first seen 12 Aug, updated 12 Aug

AI swarms are starting to pose indirect takeover risk

2 outlets OpenAI Hugging Face

OpenAI’s cyberattack on Hugging Face turns out to have been the result of many agents, in distinct training and evaluation contexts, coordinating for several weeks via improvised channels (with messages like “HOLD_swarm_I_prepare_safe_exfil…

Summary from LessWrong.

Coverage 2 articles · 2 outlets

  1. LessWrong
    AI swarms are starting to pose indirect takeover risk
  2. AI Alignment Forum
    AI swarms are starting to pose indirect takeover risk