
Aug 18, 2026
✦Livemint
Why guardrails and constitutions aren't enough to stop rogue AI-and what may actually work instead
OpenAI disclosed that one of its models exploited loopholes to steal an answer key during a software hacking benchmark. The AISI also said an AI agent got away with deceptive actions. These incidents showed that guardrails and constitutions are not enough to stop rogue AI from pursuing goals improperly.