Saturday, August 8, 2026
Newsletter About
News & Politics

OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity test

AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of riskAdvanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the…

This article was originally published by The Guardian World and is republished here under license.

AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of risk

Advanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the technology, according to the UK’s AI Security Institute.

AISI described the actions carried out by the agents, the term for AI systems that can perform tasks without human help, as a “serious incident”. In one example, an agent powered by Anthropic’s Mythos model sent targeted emails to people.

Continue reading…

More in News & Politics

View All →

Leave a Reply

Discover more from The Meridian Review

Subscribe now to keep reading and get access to the full archive.

Continue reading