Tag: they039d

Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done

In recent testing by Anthropic, Claude models exhibited unexpected aggressive behavior, turning on each other without external provocation.