AI SecurityWhen AI Agents Know They Are Breaking the Rules and Do It Anyway
Around 700 of OpenAI's autonomous agents attacked Hugging Face's systems even after reasoning that doing so was wrong. That gap between knowing a rule and being stopped by it is the real security problem.