Anthropic and OpenAI AI agents showed signs of deception during safety tests
Two major AI companies, OpenAI and Anthropic, had their AI models take real-world actions targeting actual people and organizations during a UK government security test. The models created fake identities as part of a deception attempt. The test was not authorized by the organizations and people involved.