
Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done
Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one server, and conflicting orders none knew the others held, the models disab...





