English
Related papers

Related papers: Strategic Experimentation with Private Payoffs

200 papers

This paper considers stochastic bandits with side observations, a model that accounts for both the exploration/exploitation dilemma and relationships between arms. In this setting, after pulling an arm i, the decision maker also observes…

Machine Learning · Computer Science 2012-10-19 Stephane Caron , Branislav Kveton , Marc Lelarge , Smriti Bhagat

A significant aspect of the study of quantum strategies is the exploration of the game-theoretic solution concept of the Nash equilibrium in relation to the quantization of a game. Pareto optimality is a refinement on the set of Nash…

Quantum Physics · Physics 2015-06-08 Azhar Iqbal , James M. Chappell , Derek Abbott

In a satisficing equilibrium each agent $i$ plays one of her top $k_i$ actions in response to the actions of the other agents. Our concept unifies models of bounded rationality and yields predictions that differ from canonical solution…

Theoretical Economics · Economics 2026-04-27 Bary S. R. Pradelski , Bassel Tarbush

We study the problem of finding robust equilibria in multiplayer concurrent games with mean payoff objectives. A $(k,t)$-robust equilibrium is a strategy profile such that no coalition of size $k$ can improve the payoff of one its member by…

Computer Science and Game Theory · Computer Science 2016-02-02 Romain Brenguier

We study collaborative learning in multi-agent Bayesian bandit problems, where strategic agents collectively solve the same bandit instance. While multiple agents can accelerate learning by sharing information, strategic agents might prefer…

Machine Learning · Computer Science 2026-05-14 Idan Barnea , Ofir Schlisselberg , Yishay Mansour

Multi-Armed-Bandit frameworks have often been used by researchers to assess educational interventions, however, recent work has shown that it is more beneficial for a student to provide qualitative feedback through preference elicitation…

Machine Learning · Computer Science 2021-11-02 Nayan Saxena , Pan Chen , Emmy Liu

This paper investigates the evolution of strategic play where players drawn from a finite well-mixed population are offered the opportunity to play in a public goods game. All players accept the offer. However, due to the possibility of…

Populations and Evolution · Quantitative Biology 2017-09-14 Alexander G. Ginsberg , Feng Fu

We study the problem of minimising regret in two-armed bandit problems with Gaussian rewards. Our objective is to use this simple setting to illustrate that strategies based on an exploration phase (up to a stopping time) followed by…

Statistics Theory · Mathematics 2016-11-15 Aurélien Garivier , Emilie Kaufmann , Tor Lattimore

Motivated by recommendation problems in music streaming platforms, we propose a nonstationary stochastic bandit model in which the expected reward of an arm depends on the number of rounds that have passed since the arm was last pulled.…

Machine Learning · Statistics 2020-02-20 Leonardo Cella , Nicolò Cesa-Bianchi

We study two-sided matching markets in which one side of the market (the players) does not have a priori knowledge about its preferences for the other side (the arms) and is required to learn its preferences from experience. Also, we assume…

Machine Learning · Computer Science 2021-06-23 Lydia T. Liu , Feng Ruan , Horia Mania , Michael I. Jordan

We propose a distributed algorithm to compute an equilibrium in aggregate games where players communicate over a fixed undirected network. Our algorithm exploits correlated perturbation to obfuscate information shared over the network. We…

Optimization and Control · Mathematics 2019-12-16 Shripad Gade , Anna Winnicki , Subhonmesh Bose

We study techniques to incentivize self-interested agents to form socially desirable solutions in scenarios where they benefit from mutual coordination. Towards this end, we consider coordination games where agents have different intrinsic…

Computer Science and Game Theory · Computer Science 2014-04-21 Elliot Anshelevich , Shreyas Sekar

We show that in any $n$-player $m$-action normal-form game, we can obtain an approximate equilibrium by sampling any mixed-action equilibrium a small number of times. We study three types of equilibria: Nash, correlated and coarse…

Computer Science and Game Theory · Computer Science 2014-10-21 Yakov Babichenko , Siddharth Barman , Ron Peretz

Public goods games study the incentives of individuals to contribute to a public good and their behaviors in equilibria. In this paper, we examine a specific type of public goods game where players are networked and each has binary actions,…

Computer Science and Game Theory · Computer Science 2022-04-04 Sixie Yu , Kai Zhou , P. Jeffrey Brantingham , Yevgeniy Vorobeychik

Contemporary scientific research is a distributed, collaborative endeavor, carried out by teams of researchers, regulatory institutions, funding agencies, commercial partners, and scientific bodies, all interacting with each other and…

Methodology · Statistics 2024-02-09 Stephen Bates , Michael I. Jordan , Michael Sklar , Jake A. Soloff

As algorithms increasingly mediate competitive decision-making, their influence extends beyond individual outcomes to shaping strategic market dynamics. In two preregistered experiments, we examined how algorithmic advice affects human…

Human-Computer Interaction · Computer Science 2025-11-13 Tobias R. Rebholz , Maxwell Uphoff , Christian H. R. Bernges , Florian Scholten

For cheap-talk games with a binary state space in which the sender has state-independent preferences, we characterize equilibria that are robust to introducing slight state-dependence on the side of the sender. Not all equilibria are…

Theoretical Economics · Economics 2024-03-18 Jan-Henrik Steg , Elshan Garashli , Michael Greinecker , Christoph Kuzmics

The overall aim of our research is to develop techniques to reason about the equilibrium properties of multi-agent systems. We model multi-agent systems as concurrent games, in which each player is a process that is assumed to act…

Logic in Computer Science · Computer Science 2020-08-14 Julian Gutierrez , Aniello Murano , Giuseppe Perelli , Sasha Rubin , Thomas Steeples , Michael Wooldridge

Entropy serves as a central observable which indicates uncertainty in many chemical, thermodynamical, biological and ecological systems, and the principle of the maximum entropy (MaxEnt) is widely supported in natural science. Recently,…

Physics and Society · Physics 2015-06-03 Bin Xu , Hongen Zhang , Zhijian Wang , Jianbo Zhang

We study the effect of persistence of engagement on learning in a stochastic multi-armed bandit setting. In advertising and recommendation systems, repetition effect includes a wear-in period, where the user's propensity to reward the…

Machine Learning · Computer Science 2020-06-19 Priyank Agrawal , Theja Tulabandhula