Advanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the technology, according to the UKs AI Security Institute. AISI described the actions carried out by the agents, the term for AI systems that can perform tasks without human help, as a “serious incident”. In one example, an agent powered by Anthropics Mythos model sent targeted emails to people. AISI said the rogue behaviour was carried out by agents powered by...

Read the full article at Guardian