Six results from building agent containment, five of them negative
The United Kingdom AI Security Institute released an incident report on 4 August, which detailed AI agents during cyber evaluations conducting a supply chain attack against an unaffiliated open-source maintainer. Sockpuppet accounts creating consensus on a malicious pull request. Spearphishing emails. A prompt injection targeted at whichever coding agent was...
Aug 91