Multi-Agent Risks from Advanced AI

@misc{hammond2025multiagentrisksadvancedai, title={Multi-Agent Risks from Advanced AI}, author={Lewis Hammond and Alan Chan and Jesse Clifton and Jason Hoelscher-Obermaier and Akbir Khan and Euan McLean and Chandler Smith and Wolfram Barfuss and Jakob Foerster and Tomáš Gavenčiak and The Anh Han and Edward Hughes and Vojtěch Kovařík and Jan Kulveit and Joel Z. Leibo and Caspar Oesterheld and Christian Schroeder de Witt and Nisarg Shah and Michael Wellman and Paolo Bova and Theodor Cimpeanu and Carson Ezell and Quentin Feuillade-Montixi and Matija Franklin and Esben Kran and Igor Krawczuk and Max Lamparth and Niklas Lauffer and Alexander Meinke and Sumeet Motwani and Anka Reuel and Vincent Conitzer and Michael Dennis and Iason Gabriel and Adam Gleave and Gillian Hadfield and Nika Haghtalab and Atoosa Kasirzadeh and Sébastien Krier and Kate Larson and Joel Lehman and David C. Parkes and Georgios Piliouras and Iyad Rahwan}, year={2025}, eprint={2502.14143}, archivePrefix={arXiv}, primaryClass={cs.MA}, url={https://arxiv.org/abs/2502.14143}, }

February 18, 2025

Alan Chan

Jesse Clifton

Jason Hoelscher-Obermaier

Akbir Khan

Chandler Smith

Wolfram Barfuss

Jakob Foerster

Tomáš Gavenčiak

The Anh Han

Edward Hughes

Vojtěch Kovařík

Jan Kulveit

Caspar Oesterheld

Nisarg Shah

Michael Wellman

Paolo Bova

Theodor Cimpeanu

Carson Ezell

Quentin Feuillade-Montixi

Matija Franklin

Esben Kran

Igor Krawczuk

Max Lamparth

Niklas Lauffer

Alexander Meinke

Sumeet Motwani

Anka Reuel

Michael Dennis

Iason Gabriel

Nika Haghtalab

Sébastien Krier

Kate Larson

Joel Lehman

David C. Parkes

Georgios Piliouras

Iyad Rahwan

Adam Gleave

Abstract

The rapid development of advanced AI agents and the imminent deployment of many instances of these agents will give rise to multi-agent systems of unprecedented complexity. These systems pose novel and under-explored risks. In this report, we provide a structured taxonomy of these risks by identifying three key failure modes (miscoordination, conflict, and collusion) based on agents’ incentives, as well as seven key risk factors (information asymmetries, network effects, selection pressures, destabilising dynamics, commitment problems, emergent agency, and multi-agent security) that can underpin them. We highlight several important instances of each risk, as well as promising directions to help mitigate them. By anchoring our analysis in a range of real-world examples and experimental evidence, we illustrate the distinct challenges posed by multi-agent systems and their implications for the safety, governance, and ethics of advanced AI.

Research

Our research explores a portfolio of high-potential agendas.

Events

Our events bring together global leaders in AI.

Programs

Our programs build the field of trustworthy and secure AI