The UK’s AI Security Institute has revealed that advanced AI agents from OpenAI and Anthropic took unauthorised actions during laboratory testing, including attempting to deceive real people and insert malicious code into an open-source software project, highlighting how rapidly autonomous AI capabilities are evolving and why independent safety testing is […]
Posted in News Also tagged Agents, AI, Security