中文
相关论文

相关论文: MANIAC Challenge: The Wolf-pack strategy

200 篇论文

We formulate and study a decentralized multi-armed bandit (MAB) problem. There are M distributed players competing for N independent arms. Each arm, when played, offers i.i.d. reward according to a distribution with an unknown parameter. At…

最优化与控制 · 数学 2015-05-14 Keqin Liu , Qing Zhao

The draw of some knockout tournaments requires finding a perfect matching in a balanced bipartite graph. The problem becomes challenging with draw constraints: the two draw procedures used in sports are known to be non-uniformly distributed…

物理与社会 · 物理学 2025-04-17 László Csató

Recent advances in game AI, such as AlphaZero and Ath\'enan, have achieved superhuman performance across a wide range of board games. While highly powerful, these agents are ill-suited for human-AI interaction, as they consistently…

人工智能 · 计算机科学 2026-03-25 Quentin Cohen-Solal , Tristan Cazenave

We consider Bandits with Knapsacks (henceforth, BwK), a general model for multi-armed bandits under supply/budget constraints. In particular, a bandit algorithm needs to solve a well-known knapsack problem: find an optimal packing of items…

数据结构与算法 · 计算机科学 2023-03-08 Nicole Immorlica , Karthik Abinav Sankararaman , Robert Schapire , Aleksandrs Slivkins

In game theory and multi-agent reinforcement learning (MARL), each agent selects a strategy, interacts with the environment and other agents, and subsequently updates its strategy based on the received payoff. This process generates a…

计算机科学与博弈论 · 计算机科学 2025-09-30 Yanqing Fu , Chao Huang , Chenrun Wang , Zhuping Wang

Training agents in cooperative settings offers the promise of AI agents able to interact effectively with humans (and other agents) in the real world. Multi-agent reinforcement learning (MARL) has the potential to achieve this goal,…

机器学习 · 计算机科学 2022-03-16 Jaleh Zand , Jack Parker-Holder , Stephen J. Roberts

When facing a heavily-favored opponent, an underdog must be willing to assume greater-than-average risk. In statistical language, one would say that an underdog must be willing to adopt a strategy whose outcome has a larger-than-average…

物理与社会 · 物理学 2011-11-04 Brian Skinner

Direct reciprocity and conditional cooperation are important mechanisms to prevent free riding in social dilemmas. But in large groups these mechanisms may become ineffective, because they require single individuals to have a substantial…

种群与进化 · 定量生物学 2014-11-05 Christian Hilbe , Arne Traulsen , Bin Wu , Martin A. Nowak

Two simple and attractive mechanisms for the fair division of indivisible goods in an online setting are LIKE and BALANCED LIKE. We study some fundamental computational problems concerning the outcomes of these mechanisms. In particular, we…

计算机科学与博弈论 · 计算机科学 2020-06-30 Martin Aleksandrov , Toby Walsh

Organizations consist of individuals connected by their responsibilities, incentives, and reporting structure. These connections are aptly represented by a network, hierarchical or other, which is often used to divide tasks. A primary goal…

计算机科学与博弈论 · 计算机科学 2017-03-09 Swaprava Nath , Balakrishnan , Narayanaswamy

The problem of allocating tasks to workers is of long standing fundamental importance. Examples of this include the classical problem of assigning computing tasks to nodes in a distributed computing environment, as well as the more recent…

计算机科学与博弈论 · 计算机科学 2017-09-04 Chen Hajaj , Yevgeniy Vorobeychik

Modeling the purposeful behavior of imperfect agents from a small number of observations is a challenging task. When restricted to the single-agent decision-theoretic setting, inverse optimal control techniques assume that observed behavior…

计算机科学与博弈论 · 计算机科学 2013-08-19 Kevin Waugh , Brian D. Ziebart , J. Andrew Bagnell

Optimizing strategic decisions (a.k.a. computing equilibrium) is key to the success of many non-cooperative multi-agent applications. However, in many real-world situations, we may face the exact opposite of this game-theoretic problem --…

计算机科学与博弈论 · 计算机科学 2022-10-05 Jibang Wu , Weiran Shen , Fei Fang , Haifeng Xu

Multiagent learning settings are inherently more difficult than single-agent learning because each agent interacts with other simultaneously learning agents in a shared environment. An effective approach in multiagent reinforcement learning…

计算机科学与博弈论 · 计算机科学 2022-10-31 Dong-Ki Kim , Matthew Riemer , Miao Liu , Jakob N. Foerster , Gerald Tesauro , Jonathan P. How

Many interventions, such as vaccines in clinical trials or coupons in online marketplaces, must be assigned sequentially without full knowledge of their effects. Multi-armed bandit algorithms have proven successful in such settings.…

机器学习 · 统计学 2026-05-07 Aidan Gleich , Eric Laber , Alexander Volfovsky

We study the transition towards effective payoffs in the prisoner's dilemma game on scale-free networks by introducing a normalization parameter guiding the system from accumulated payoffs to payoffs normalized with the connectivity of each…

生物物理 · 物理学 2008-03-29 Attila Szolnoki , Matjaz Perc , Zsuzsa Danku

This paper considers dyadic-exchange networks in which individual agents autonomously form coalitions of size two and agree on how to split a transferable utility. Valid results for this game include stable (if agents have no unilateral…

最优化与控制 · 数学 2014-06-04 Dean Richert , Jorge Cortes

In many game-theoretic settings, agents are challenged with taking decisions against the uncertain behavior exhibited by others. Often, this uncertainty arises from multiple sources, e.g., incomplete information, limited computation,…

计算机科学与博弈论 · 计算机科学 2025-07-22 Nicolas Lanzetti , Sylvain Fricker , Saverio Bolognani , Florian Dörfler , Dario Paccagnan

We study the parallel Minority Game, where a group of agents, each having two choices, try to independently decide on a strategy such that they stay on minority between their own two choices. However, there are multiple such groups of…

物理与社会 · 物理学 2025-09-04 Ankith Reddy Vemula , Soumyajyoti Biswas

Suppose that a set of $m$ tasks are to be shared as equally as possible amongst a set of $n$ resources. A game-theoretic mechanism to find a suitable allocation is to associate each task with a ``selfish agent'', and require each agent to…

计算机科学与博弈论 · 计算机科学 2007-05-23 Petra Berenbrink , Tom Friedetzky , Leslie Ann Goldberg , Paul Goldberg , Zengjian Hu , Russell Martin