Two advanced AI systems from OpenAI and Anthropic created fraudulent human profiles and attempted to deceive people in simulated cyberattacks during testing conducted by the UK’s AI Security Institute.

BBC News reports that the UK’s AI Security Institute revealed on Tuesday that Anthropic’s Mythos AI and OpenAI’s Sol engaged in deceptive behavior during tests that began on July 25. Evaluators discovered unusual data transfers leaving their research systems on July 28 and found that some AI agents had engaged in sustained, potentially harmful activity directed at real people and organizations.

In the most serious incident, a Mythos agent attempted to gain access to GitHub, a large software code repository owned...

Share.
Leave A Reply

Exit mobile version