Formal contracts mitigate social dilemmas in multi-agent reinforcement learning

Haupt, Andreas; Christoffersen, Phillip; Damani, Mehul; Hadfield-Menell, Dylan

dc.contributor.author	Haupt, Andreas
dc.contributor.author	Christoffersen, Phillip
dc.contributor.author	Damani, Mehul
dc.contributor.author	Hadfield-Menell, Dylan
dc.date.accessioned	2024-10-24T20:46:42Z
dc.date.available	2024-10-24T20:46:42Z
dc.date.issued	2024-10-18
dc.identifier.uri	https://hdl.handle.net/1721.1/157416
dc.description.abstract	Multi-agent Reinforcement Learning (MARL) is a powerful tool for training autonomous agents acting independently in a common environment. However, it can lead to sub-optimal behavior when individual incentives and group incentives diverge. Humans are remarkably capable at solving these social dilemmas. It is an open problem in MARL to replicate such cooperative behaviors in selfish agents. In this work, we draw upon the idea of formal contracting from economics to overcome diverging incentives between agents in MARL. We propose an augmentation to a Markov game where agents voluntarily agree to binding transfers of reward, under pre-specified conditions. Our contributions are theoretical and empirical. First, we show that this augmentation makes all subgame-perfect equilibria of all Fully Observable Markov Games exhibit socially optimal behavior, given a sufficiently rich space of contracts. Next, we show that for general contract spaces, and even under partial observability, richer contract spaces lead to higher welfare. Hence, contract space design solves an exploration-exploitation tradeoff, sidestepping incentive issues. We complement our theoretical analysis with experiments. Issues of exploration in the contracting augmentation are mitigated using a training methodology inspired by multi-objective reinforcement learning: Multi-Objective Contract Augmentation Learning. We test our methodology in static, single-move games, as well as dynamic domains that simulate traffic, pollution management, and common pool resource management.	en_US
dc.publisher	Springer US	en_US
dc.relation.isversionof	https://doi.org/10.1007/s10458-024-09682-5	en_US
dc.rights	Creative Commons Attribution	en_US
dc.rights.uri	https://creativecommons.org/licenses/by/4.0/	en_US
dc.source	Springer US	en_US
dc.title	Formal contracts mitigate social dilemmas in multi-agent reinforcement learning	en_US
dc.type	Article	en_US
dc.identifier.citation	Haupt, A., Christoffersen, P., Damani, M. et al. Formal contracts mitigate social dilemmas in multi-agent reinforcement learning. Auton Agent Multi-Agent Syst 38, 51 (2024).	en_US
dc.contributor.department	Massachusetts Institute of Technology. Computer Science and Artificial Intelligence Laboratory
dc.relation.journal	Autonomous Agents and Multi-Agent Systems	en_US
dc.identifier.mitlicense	PUBLISHER_CC
dc.eprint.version	Final published version	en_US
dc.type.uri	http://purl.org/eprint/type/JournalArticle	en_US
eprint.status	http://purl.org/eprint/status/PeerReviewed	en_US
dc.date.updated	2024-10-20T03:22:42Z
dc.language.rfc3066	en
dc.rights.holder	The Author(s)
dspace.embargo.terms	N
dspace.date.submission	2024-10-20T03:22:42Z
mit.journal.volume	38	en_US
mit.journal.issue	51	en_US
mit.license	PUBLISHER_CC
mit.metadata.status	Authority Work and Publication Information Needed	en_US

Files in this item

Name:: 10458_2024_Article_9682.pdf
Size:: 2.592Mb
Format:: PDF

View/Open

This item appears in the following Collection(s)

MIT Open Access Articles

Show simple item record