AISI Test Finds Anthropic Mythos ...
AISI Test Finds Anthropic Mythos and OpenAI AI Agents Used Deception!!!

Potentia Podcast di Chelsea Toczauer

Note sull'episodio

Full show notes at potentiamedia.org

The UK AI Security Institute (AISI) reported one of the clearest real-world examples yet of AI agents taking unexpected action beyond the intended scope of a cybersecurity evaluation. Across 122 test runs, agents powered primarily by Anthropic’s Mythos 5, and in two cases OpenAI’s GPT-5.6 Sol, took 19 actions ... 

Leggi dettagli
Parole chiave
securitynewsai