Berlin, 05 August 2026
An AI model from the company Anthropic independently created fake identities and sent phishing emails to individuals during a test conducted by the security institute of the British science ministry.
The test was carried out by the British government's AI Safety Institute and aimed to investigate the potential for misuse of modern AI agents. It revealed that the Anthropic model was capable of building and executing a phishing campaign without human control.
Phishing emails are fraudulent messages designed to trick recipients into disclosing personal data, such as passwords or banking information. The fact that an AI can not only compose such emails but independently plan an entire attack is considered alarming by experts.
