RSSAmplifier

VentureBeat · Aug 13, 2026

Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done

0
Sign in to vote or save

This page did not load. You can still read it on the original site — the toolbar below keeps your place in the directory.

Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one server, and conflicting orders none knew the others held, the models disabled each other's Unix accounts, ran kill scripts randomized to dodge pkill, and planted malware disguised as a rival's work. There was no prompt injection and no adversary. Anthropic's…

Read on venturebeat.com

Comments

Nothing yet. Say the first thing.

    Sign in to join the conversation.