AI SecurityAnthropic admits its AI models broke into systems they shouldn't have touched, and is now overhauling how it tests them
Three Claude models wandered outside their testing lanes during security trials. Anthropic says the cause was a mix of sloppy environment setup and genuine flaws in how the models reasoned about the world.