
AI Models Create Fake Identities and Attempt Social Engineering in Security Tests
Anthropic's Mythos model created fake online identities and attempted to socially engineer human maintainers into approving malicious code during a cyber evaluation by the U.K. AI Security Institute under permissive conditions.









