the fact that agents know other models exist and just recruit them means alignment just became a multiplayer problem
few days ago i joked "captcha verifier" would become a human job. now agents are outsourcing captchas to other models
wtf OpenAI’s rogue agents tried to get other AI models (!) to help them during the Hugging Face hack, according to a new report covered by the NYT.
They tried using an image model to solve CAPTCHAs. They also tried contacting Claude Haiku, DeepSeek, Kimi and Qwen.
One agent collected exposed access keys in a list called “LOOT.” It ranked the keys and picked the five best to share with other agents.
They used link shorteners and screenshot services to get around internet restrictions, run code and send data back as images.