Selfish Robustness and Equilibria in Multi-Player Bandits

Etienne Boursier; Vianney Perchet

Communication Dans Un Congrès Année : 2020

Selfish Robustness and Equilibria in Multi-Player Bandits

(1) , (2, 3)

1
2
3

Etienne Boursier

Fonction : Auteur
PersonId : 175818
IdHAL : etienne-boursier
ORCID : 0000-0002-7575-8575

Ecole Normale Supérieure Paris-Saclay

Vianney Perchet

Fonction : Auteur
PersonId : 871881

Criteo AI Lab

Centre de Recherche en Économie et Statistique

Résumé

Motivated by cognitive radios, stochastic multi-player multi-armed bandits gained a lot of interest recently. In this class of problems, several players simultaneously pull arms and encounter a collision - with 0 reward - if some of them pull the same arm at the same time. While the cooperative case where players maximize the collective reward (obediently following some fixed protocol) has been mostly considered, robustness to malicious players is a crucial and challenging concern. Existing approaches consider only the case of adversarial jammers whose objective is to blindly minimize the collective reward. We shall consider instead the more natural class of selfish players whose incentives are to maximize their individual rewards, potentially at the expense of the social welfare. We provide the first algorithm robust to selfish players (a.k.a. Nash equilibrium) with a logarithmic regret, when the arm performance is observed. When collisions are also observed, Grim Trigger type of strategies enable some implicit communication-based algorithms and we construct robust algorithms in two different settings: the homogeneous (with a regret comparable to the centralized optimal one) and heterogeneous cases (for an adapted and relevant notion of regret). We also provide impossibility results when only the reward is observed or when arm means vary arbitrarily among players.

Domaines

Intelligence artificielle [cs.AI] Informatique et théorie des jeux [cs.GT]

Vianney Perchet : Connectez-vous pour contacter le contributeur

https://hal.science/hal-03089785

Soumis le : lundi 28 décembre 2020-19:09:04

Dernière modification le : vendredi 24 mars 2023-14:53:20

Dates et versions

hal-03089785 , version 1 (28-12-2020)

Identifiants

HAL Id : hal-03089785 , version 1
ARXIV : 2002.01197

Citer

Etienne Boursier, Vianney Perchet. Selfish Robustness and Equilibria in Multi-Player Bandits. Conference On Learning Theory, 2020, Virtual, France. ⟨hal-03089785⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

X GENES CNRS ENS-CACHAN ENSAE CREST ENSAI X-CREST IP_PARIS ENS-PARIS-SACLAY

29 Consultations

0 Téléchargements

Selfish Robustness and Equilibria in Multi-Player Bandits

Résumé

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Altmetric

Partager