Skip to content

Google teases an ATT&CK framework for agents: Is yours a TRAIT&R?

At risk level 2, it will cost a lot more to stop your agent fleet from going traitor, warns the GDM AI Control Roadmap v0.1.

Google teases an ATT&CK framework for agents: Is yours a TRAIT&R?
Google's depiction of its AI Control Roadmap.

Google DeepMind's new security framework using the popular MITRE ATT&CK approach prepares enterprises for the future dangers of rogue AI agents. 

Google describes the v0.1 GDM AI Control Roadmap as a "first-of-its-kind blueprint for internal security against potentially misaligned AI".

Although human and AI attackers differ, the authors argue the ATT&CK approach to tactics, techniques, and procedures can be extended to what they have named TRAIT&R: a Taxonomy of Rogue AI Tactics and Routines. 

Much of the roadmap relies on forecasts of future AI capabilities.

However, it comes with the certainty that, as model capability rises, so will the cost of defence. 

This content is for paying members only

Subscribe
Add The Stack on Google