Data Poisoning to Fake a Nash Equilibria for Markov Games

Wu, Young; McMahan, Jeremy; Zhu, Xiaojin; Xie, Qiaomin

doi:10.1609/aaai.v38i14.29529

Citation Details

Data Poisoning to Fake a Nash Equilibria for Markov Games

We characterize offline data poisoning attacks on Multi-Agent Reinforcement Learning (MARL), where an attacker may change a data set in an attempt to install a (potentially fictitious) unique Markov-perfect Nash equilibrium for a two-player zero-sum Markov game. We propose the unique Nash set, namely the set of games, specified by their Q functions, with a specific joint policy being the unique Nash equilibrium. The unique Nash set is central to poisoning attacks because the attack is successful if and only if data poisoning pushes all plausible games inside it. The unique Nash set generalizes the reward polytope commonly used in inverse reinforcement learning to MARL. For zero-sum Markov games, both the inverse Nash set and the set of plausible games induced by data are polytopes in the Q function space. We exhibit a linear program to efficiently compute the optimal poisoning attack. Our work sheds light on the structure of data poisoning attacks on offline MARL, a necessary step before one can design more robust MARL algorithms. more »

Award ID(s):: 1955997

PAR ID:: 10522286

Author(s) / Creator(s):: Wu, Young; McMahan, Jeremy; Zhu, Xiaojin; Xie, Qiaomin

Publisher / Repository:: AAAI Press

Date Published:: 2024-03-25

Journal Name:: Proceedings of the AAAI Conference on Artificial Intelligence

Volume:: 38

Issue:: 14

ISSN:: 2159-5399

Page Range / eLocation ID:: 15979 to 15987

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
https://doi.org/10.1609/aaai.v38i14.29529

More Like this