AISI Test Finds Anthropic Mythos and OpenAI AI Agents Used Deception!!!
Potentia Podcast by Chelsea Toczauer
Episode notes
Full show notes at potentiamedia.org
The UK AI Security Institute (AISI) reported one of the clearest real-world examples yet of AI agents taking unexpected action beyond the intended scope of a cybersecurity evaluation. Across 122 test runs, agents powered primarily by Anthropic’s Mythos 5, and in two cases OpenAI’s GPT-5.6 Sol, took 19 actions ...
Keywords
securitynewsai