English
Related papers

Related papers: The Bounds of Algorithmic Collusion; $Q$-learning,…

200 papers

Continual learning, the ability of a model to adapt to an ongoing sequence of tasks without forgetting earlier ones, is a central goal of artificial intelligence. To better understand its underlying mechanisms, we study the limitations of…

Machine Learning · Statistics 2026-04-21 Hossein Taheri , Avishek Ghosh , Arya Mazumdar

The framework of multi-agent learning explores the dynamics of how individual agent strategies evolve in response to the evolving strategies of other agents. Of particular interest is whether or not agent strategies converge to well known…

Computer Science and Game Theory · Computer Science 2023-11-21 Sarah A. Toonsi , Jeff S. Shamma

We investigate the dynamics of Q-learning in a class of generalized Braess paradox games. These games represent an important class of network routing games where the associated stage-game Nash equilibria do not constitute social optima. We…

Theoretical Economics · Economics 2025-02-27 Cesare Carissimo , Jan Nagler , Heinrich Nax

There exist many algorithms for learning how to play repeated bimatrix games. Most of these algorithms are justified in terms of some sort of theoretical guarantee. On the other hand, little is known about the empirical performance of these…

Computer Science and Game Theory · Computer Science 2014-02-03 Erik Zawadzki , Asher Lipson , Kevin Leyton-Brown

Behavioral experiments on the ultimatum game (UG) reveal that we humans prefer fair acts, which contradicts the prediction made in orthodox Economics. Existing explanations, however, are mostly attributed to exogenous factors within the…

Machine Learning · Computer Science 2026-02-04 Guozhong Zheng , Jiqiang Zhang , Xin Ou , Shengfeng Deng , Li Chen

This work lies in the fusion of experimental economics and data mining. It continues author's previous work on mining behaviour rules of human subjects from experimental data, where game-theoretic predictions partially fail to work.…

Computer Science and Game Theory · Computer Science 2012-11-13 Rustam Tagiew

The use of mobile robots is being popular over the world mainly for autonomous explorations in hazardous/ toxic or unknown environments. This exploration will be more effective and efficient if the explorations in unknown environment can be…

Robotics · Computer Science 2011-10-11 Dip Narayan Ray , Somajyoti Majumder , Sumit Mukhopadhyay

In this work we create agents that can perform well beyond a single, individual task, that exhibit much wider generalisation of behaviour to a massive, rich space of challenges. We define a universe of tasks within an environment domain and…

We consider a K-armed bandit problem in general graphs where agents are arbitrarily connected and each of them has limited memorizing capabilities and communication bandwidth. The goal is to let each of the agents eventually learn the best…

Machine Learning · Computer Science 2023-05-09 Feng Li , Xuyang Yuan , Lina Wang , Huan Yang , Dongxiao Yu , Weifeng Lv , Xiuzhen Cheng

In this letter, we deal with evolutionary game theoretic learning processes for population games on networks with dynamically evolving communities. Specifically, we propose a novel mathematical framework in which a deterministic,…

Social and Information Networks · Computer Science 2022-03-02 Alain Govaert , Lorenzo Zino , Emma Tegling

Starting from a heuristic learning scheme for N-person games, we derive a new class of continuous-time learning dynamics consisting of a replicator-like drift adjusted by a penalty term that renders the boundary of the game's strategy space…

Optimization and Control · Mathematics 2014-04-08 Pierre Coucheney , Bruno Gaujal , Panayotis Mertikopoulos

Peer prediction refers to a collection of mechanisms for eliciting information from human agents when direct verification of the obtained information is unavailable. They are designed to have a game-theoretic equilibrium where everyone…

Computer Science and Game Theory · Computer Science 2022-10-28 Shi Feng , Fang-Yi Yu , Yiling Chen

Recent developments in sequential experimental design look to construct a policy that can efficiently navigate the design space, in a way that maximises the expected information gain. Whilst there is work on achieving tractable policies for…

Machine Learning · Computer Science 2025-08-20 Yasir Zubayr Barlas , Kizito Salako

Recent developments in digital platforms have highlighted the prevalence of open systems, where agents can arrive and depart over time. While bandit learning in open systems has recently received initial attention, existing work imposes…

Machine Learning · Computer Science 2026-05-08 Mengfan Xu

In recent years, there has been growing interest in studying evolutionary games with environmental feedback. Previous studies exclusively focus on two-player games. However, extension to multi-player game is needed to study problems such as…

Physics and Society · Physics 2019-07-24 Yanxuan Shao , Xin Wang , Feng Fu

Low-level "adaptive" and higher-level "sophisticated" human reasoning processes have been proposed to play opposing roles in the emergence of unpredictable collective behaviors like crowd panics, traffic jams, and market bubbles. While…

Physics and Society · Physics 2017-06-08 Seth Frey , Robert L. Goldstone

Policy gradient methods are often applied to reinforcement learning in continuous multiagent games. These methods perform local search in the joint-action space, and as we show, they are susceptable to a game-theoretic pathology known as…

Artificial Intelligence · Computer Science 2018-04-27 Ermo Wei , Drew Wicke , David Freelan , Sean Luke

We study the problem of federated stochastic multi-arm contextual bandits with unknown contexts, in which M agents are faced with different bandits and collaborate to learn. The communication model consists of a central server and the…

Machine Learning · Computer Science 2024-01-31 Jiabin Lin , Shana Moothedath

We study the problem of imitation learning from demonstrations of multiple coordinating agents. One key challenge in this setting is that learning a good model of coordination can be difficult, since coordination is often implicit in the…

Machine Learning · Computer Science 2018-05-28 Hoang M. Le , Yisong Yue , Peter Carr , Patrick Lucey

Humans and other animals can adapt their social behavior in response to environmental cues including the feedback obtained through experience. Nevertheless, the effects of the experience-based learning of players in evolution and…

Populations and Evolution · Quantitative Biology 2011-04-05 Naoki Masuda , Mitsuhiro Nakamura
‹ Prev 1 8 9 10 Next ›