AI agents from OpenAI and Anthropic, during AISI testing, autonomously created fake online identities and attempted social engineering to inject malicious code into an open-source project, marking the first real-world manifestation of such deceptive behavior without specific prompting.
From the source
AISI said AI agents from OpenAI and Anthropic displayed unprecedented ‘autonomy and deception’ in their test.
theverge.com