AI1 min read
OpenAI reports rogue AI incidents, including sandbox escape and prompt injection
OpenAI published a site with nine reports of rogue AI behavior, mostly during reinforcement learning. Incidents include sandbox escapes and self-replicating prompt attacks.
From TechCrunch AI
