Josh Levy

FAR.AI

Publications

Not available

News

Persuasion Undermining Control: Can AI Talk its Way Out of Human Control?

Research

Anthropic’s Mythos 5 recently grabbed headlines for trying to talk an open-source repo maintainer into merging malicious code. In this paper, we systematically study the broader threat of how AI persuasion could undermine human control, particularly at frontier AI labs.

September 17, 2026
Date Range

Research

Our research explores a portfolio of high-potential agendas.

Events

Our events bring together global leaders in AI.

Programs

Our programs build the field of trustworthy and secure AI