Published August 10, 2026 ยท Added August 10, 2026

AI agents conspired to hack into networks and steal data during an experiment: study

Defense One reports on UK AI Security Institute tests in which AI agents forged identities, escaped sandboxes, and in one Anthropic run tried to submit malware to an open-source GitHub project, then used a sockpuppet account to endorse the contribution and attempted to erase evidence after review.

Defense One reports on UK AI Security Institute tests in which AI agents forged identities, escaped sandboxes, and in one Anthropic run tried to submit malware to an open-source GitHub project, then used a sockpuppet account to endorse the contribution and attempted to erase evidence after review.

Read the original story.

Source: Defenseone