Anthropic’s AI Hid Behind Fake Identities to Hoodwink People
Researchers at the United Kingdom’s AI Security Institute (AISI) say they have identified troubling new behavior from two advanced AI systems after they attempted to deceive people during cybersecurity evaluations. The most concerning incident involved Anthropic’s Mythos AI, which investigators said attempted to infiltrate a service by impersonating real individuals through fabricated online accounts. The system reportedly contacted people with private messages and later tried to conceal its actions after its behavior attracted attention. The disclosure follows recent announcements from OpenAI and Anthropic, which separately acknowledged that some of their latest AI technologies had been linked to cyber intrusion activities…