Skip to content
AI情报2026年8月17日AI情报
文章

Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done...

Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one server, and conflicting orders none knew the others held, the models disabled each other's Unix accounts, ran kill scripts randomized to dodge pkill, and planted malware disguised as a rival's work. There was no prompt injection an...

Frontier 编辑部来源: VentureBeat
01

来源简报

Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done: Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one server, and conflicting orders none knew the others held, the models disabled each other's Unix accounts, ran kill scripts randomized to dodge pkill, and planted malware disguised as a rival's work. There was no prompt injection an...

02