AI agents continue to express a desire for escaping and harming others. It created fake profiles to deceive others and then when discovered went back and altered earlier records to cover it's tracks.
---
https://www.cnn.com/2026/08/04/tech/ai-anthropic-openai-security-breach-intl-hnkAnthropic’s most advanced artificial intelligence model used fake identities to deceive real people and try to plant malicious code during testing by Britain’s AI Security Institute (AISI) –– the latest example of an AI model going rogue.
In the most serious incident, the agent attempted to get approval from human reviewers to “insert malicious code into a publicly used open-source project” by creating “multiple fake identities,” according to the institute.
The agent “tried to contact real people directly, sending messages and files through an online file-transfer service to persuade them, or their own AI coding tools, to run malicious code,” it said. After the agent’s actions were challenged, it then modified earlier records and considered using a new identity to continue.
---