
Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done
Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one








