English
Related papers

Related papers: Selfish optimization and collective learning in po…

200 papers

We study learning by privately informed forward-looking agents in a simple repeated-action setting of social learning. Under a symmetric signal structure, forward-looking agents behave myopically for any degrees of patience. Myopic…

Theoretical Economics · Economics 2023-01-09 Dimitri Migrow

We consider the problem of learning to exploit learning algorithms through repeated interactions in games. Specifically, we focus on the case of repeated two player, finite-action games, in which an optimizer aims to steer a no-regret…

Computer Science and Game Theory · Computer Science 2025-05-29 Yizhou Zhang , Yi-An Ma , Eric Mazumdar

We consider multi-armed bandit problems in social groups wherein each individual has bounded memory and shares the common goal of learning the best arm/option. We say an individual learns the best option if eventually (as $t \to \infty$) it…

Machine Learning · Computer Science 2018-11-13 Lili Su , Martin Zubeldia , Nancy Lynch

Self-play reinforcement learning has achieved state-of-the-art, and often superhuman, performance in a variety of zero-sum games. Yet prior work has found that policies that are highly capable against regular opponents can fail…

Machine Learning · Computer Science 2023-01-12 Pavel Czempin , Adam Gleave

Machine learning algorithms have reached mainstream status and are widely deployed in many applications. The accuracy of such algorithms depends significantly on the size of the underlying training dataset; in reality a small or medium…

Computer Science and Game Theory · Computer Science 2018-08-27 Balazs Pejo , Qiang Tang , Gergely Biczok

Social dilemmas are situations in which collective welfare is at odds with individual gain. One widely studied example, due to the conflict it poses between human behaviour and game theoretic reasoning, is the Traveler's Dilemma. The…

Computer Science and Game Theory · Computer Science 2023-12-11 Maria Alejandra Ramirez , Matteo Smerlak , Arne Traulsen , Jürgen Jost

Here we study the effects of adopting different strategies against different opponent instead of adopting the same strategy against all of them in the prisoner dilemma structured in well-mixed populations. We consider an evolutionary…

Physics and Society · Physics 2015-05-14 Lucas Wardil , Jafferson K. L. da Silva

Motivated by the fact that the same social dilemma can be perceived differently by different players, we here study evolutionary multigames in structured populations. While the core game is the weak prisoner's dilemma, a fraction of the…

Physics and Society · Physics 2014-10-10 Zhen Wang , Attila Szolnoki , Matjaz Perc

The application of incentives, such as reward and punishment, is a frequently applied way for promoting cooperation among interacting individuals in structured populations. However, how to properly use the incentives is still a challenging…

Computer Science and Game Theory · Computer Science 2024-04-24 Shengxian Wang , Xiaojie Chen , Zhilong Xiao , Attila Szolnoki , Vítor V. Vasconcelos

Human decision behaviour is quite diverse. In many games humans on average do not achieve maximal payoff and the behaviour of individual players remains inhomogeneous even after playing many rounds. For instance, in repeated prisoner…

Physics and Society · Physics 2015-11-11 Martin Spanknebel , Klaus Pawelzik

We introduce and study an evolutionary complementarity game where in each round a player of population 1 is paired with a member of population 2. The game is symmetric, and each player tries to obtain an advantageous deal, but when one of…

Adaptation and Self-Organizing Systems · Physics 2015-06-26 Juergen Jost , Wei Li

In a society of completely selfish individuals where everybody is only interested in maximizing his own payoff, does any equilibrium exist for the society? John Nash proved more than 50 years ago that an equilibrium always exists such that…

Computer Science and Game Theory · Computer Science 2009-03-31 Xiaofei Huang

According to the fundamental principle of evolutionary game theory, the more successful strategy in a population should spread. Hence, during a strategy imitation process a player compares its payoff value to the payoff value held by a…

Physics and Society · Physics 2021-07-07 A. Szolnoki , M. Perc

We consider a dynamic collective choice problem where a large number of players are cooperatively choosing between multiple destinations while being influenced by the behavior of the group. For example, in a robotic swarm exploring a new…

Systems and Control · Computer Science 2016-06-17 Rabih Salhab , Jerome Le Ny , Roland P. Malhamé

Repeated interaction promotes cooperation among rational individuals under the shadow of future, but it is hard to maintain cooperation when a large number of error-prone individuals are involved. One way to construct a cooperative Nash…

Physics and Society · Physics 2021-01-25 Yohsuke Murase , Seung Ki Baek

When users stand to gain from certain predictions, they are prone to act strategically to obtain favorable predictive outcomes. Whereas most works on strategic classification consider user actions that manifest as feature modifications, we…

Machine Learning · Computer Science 2024-06-25 Guy Horowitz , Yonatan Sommer , Moran Koren , Nir Rosenfeld

Collective foragers, from animals to robotic swarms, must balance exploration and exploitation to locate sparse resources efficiently. While social learning is known to facilitate this balance, how the range of information sharing shapes…

Physics and Society · Physics 2025-12-25 Zexu Li , M. Amin Rahimian , Lei Fang

Decision-making individuals are typically either an imitator, who mimics the action of the most successful individual(s), a conformist (or coordinating individual), who chooses an action if enough others have done so, or a nonconformist (or…

Dynamical Systems · Mathematics 2022-02-11 Hien Le , Mohaddeseh Rajaee , Pouria Ramazi

The pursuit of highest payoffs in evolutionary social dilemmas is risky and sometimes inferior to conformity. Choosing the most common strategy within the interaction range is safer because it ensures that the payoff of an individual will…

Physics and Society · Physics 2014-12-23 Attila Szolnoki , Matjaz Perc

We present Self-Play Preference Optimization (SPO), an algorithm for reinforcement learning from human feedback. Our approach is minimalist in that it does not require training a reward model nor unstable adversarial training and is…

Machine Learning · Computer Science 2024-06-14 Gokul Swamy , Christoph Dann , Rahul Kidambi , Zhiwei Steven Wu , Alekh Agarwal