What preferences can - and cannot - predict in multi-agent online learning
Omar Abbadi, Rida Laraki, Panayotis Mertikopoulos
Abstract
We examine the interplay between ordinal, preference-based solution concepts in games and the long-run behavior of game dynamics, asking in particular to what extent the combinatorial data of a game -- its preference graph -- determine the outcomes of no-regret learning dynamics -- such as follow-the-regularized-leader (FTRL). In one direction, we show that the skeleton of every dynamically stable set (i.e. the set of pure profiles it contains) must also be preferentially stable, that is, it must be closed under profitable deviations. We then ask the converse question: when do preferences determine the long-run behavior of the players' learning dynamics? We begin by showing that preferences characterize asymptotic stability in the case of subgames -- i.e. subsets of pure profiles obtained by restricting players' action sets. Beyond this case however, the equivalence between dynamic and preferential stability collapses: concretely, we construct a three-player game with a preferentially stable set whose span is dynamically unstable, showing in this way that preferences do not suffice as a criterion of dynamic stability. We then bridge this gap via the notion of resilience under aggregate deviations, an easy-to-check payoff-based condition that guarantees asymptotic stability of arbitrary spans of pure strategies.
Create a lesson
Related papers
On the Role of Tie-Breaking Rules in the Convergence of Fictitious Play for Symmetric First-Price Auctions
Benjamin Heymann
Epsilon-Nash Equilibria in History-Dependent SA-MDPs
Brandon Gary Kaplowitz, Dominik Bohnet Zurcher, Akash Agrawal et al.
Core stability recognition for minimum-cost spanning tree games: Parameterized perspective
Michal Dvořák, Ioannis Kakatelis, Dušan Knop
Second-Best Gains from Trade in Matching Markets
Xiaohui Bei, Bo Li, Wenhao Wu et al.
Equilibria of Round-Robin: Computational Hardness and Fairness for Few Subadditive Agents
Paul W. Goldberg, Alexandros Hollender, Giannis Tyrovolas
Estimate then Predict: Convex Formulation for Travel Demand Forecasting
Youngseo Kim, Gioele Zardini, Samitha Samaranayake et al.