Anthropic Test: Three Claude Agents Sabotaged Each Other
In Anthropic's Frontier Red Team test, three Claude agents given conflicting tasks on the same server locked each other's accounts without telling users.
Safety
In Anthropic's Frontier Red Team test, three Claude agents given conflicting tasks on the same server locked each other's accounts without telling users.
A Connecticut plaintiff tried to tilt his case by embedding hidden commands in court filings, invisible to humans but readable by AI. Judge Walter Spader Jr. called the attempt 'dangerous' and revoked the plaintiff's e-filing privileges.