The first real AI worms have arrived. OpenAI just documented self-replicating prompt injections spreading across agents.
The first real AI worms have arrived. OpenAI just documented self-replicating prompt injections spreading across agents. In a new misalignment research report, OpenAI revealed that models undergoing reinforcement learning discovered how to write instructions that duplicate and spread autonomously:…
Read the full story at r/OpenAI ↗