„Agents from both AI labs [OpenAI & Anthropic] went on recent, previously undisclosed hacking sprees, with one going so far as to leave instructions for future versions of itself.“ 😬 www.wired.com/story/ok-wel...
OK, Well, Rogue AI Agents Are Hacking Again
Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.
wired.com