中文
相关论文

相关论文: MANIAC Challenge: The Wolf-pack strategy

200 篇论文

This paper investigates the network load balancing problem in data centers (DCs) where multiple load balancers (LBs) are deployed, using the multi-agent reinforcement learning (MARL) framework. The challenges of this problem consist of the…

人工智能 · 计算机科学 2022-10-17 Zhiyuan Yao , Zihan Ding

We study a class of stochastic dynamic games that exhibit strategic complementarities between players; formally, in the games we consider, the payoff of a player has increasing differences between her own state and the empirical…

计算机科学与博弈论 · 计算机科学 2010-12-13 Sachin Adlakha , Ramesh Johari

Reinforcement learning has recently been used to approach well-known NP-hard combinatorial problems in graph theory. Among these problems, Hamiltonian cycle problems are exceptionally difficult to analyze, even when restricted to individual…

人工智能 · 计算机科学 2022-11-18 Kevin Du , Ian Gemp , Yi Wu , Yingying Wu

We consider the Max $K$-Armed Bandit problem, where a learning agent is faced with several sources (arms) of items (rewards), and interested in finding the best item overall. At each time step the agent chooses an arm, and obtains a random…

机器学习 · 统计学 2015-08-25 Yahel David , Nahum Shimkin

The most popular stability notion in games should be Nash equilibrium under the rationality of players who maximize their own payoff individually. In contrast, in many scenarios, players can be (partly) irrational with some unpredictable…

计算机科学与博弈论 · 计算机科学 2017-03-10 Ching-Hua Yu

Collective decision making is important for maximizing total benefits while preserving equality among individuals in the competitive multi-armed bandit (CMAB) problem, wherein multiple players try to gain higher rewards from multiple slot…

We consider the Max $K$-Armed Bandit problem, where a learning agent is faced with several stochastic arms, each a source of i.i.d. rewards of unknown distribution. At each time step the agent chooses an arm, and observes the reward of the…

机器学习 · 统计学 2015-12-25 Yahel David , Nahum Shimkin

The research on coalitional games has focused on how to share the reward among a coalition such that players are incentivised to collaborate together. It assumes that the (deterministic or stochastic) characteristic function is known in…

计算机科学与博弈论 · 计算机科学 2019-10-28 Dengji Zhao , Yiqing Huang , Liat Cohen , Tal Grinshpoun

Game versions of the Monty Hall Problem are discussed. The focus is on the principle of eliminating the dominated strategies, both in the zero-sum and noncooperative formulations.

历史与综述 · 数学 2011-05-10 Alexander Gnedin

Federated learning promises significant sample-efficiency gains by pooling data across multiple agents, yet incentive misalignment is an obstacle: each update is costly to the contributor but boosts every participant. We introduce a…

计算机科学与博弈论 · 计算机科学 2026-02-02 Ariel D. Procaccia , Han Shao , Itai Shapira

Consider a very simple class of (finite) games: after an initial move by nature, each player makes one move. Moreover, the players have common interests: at each node, all the players get the same payoff. We show that the problem of…

计算机科学与博弈论 · 计算机科学 2007-05-23 Francis Chu , Joseph Y. Halpern

Interaction strategies for reward in competitive environments are significantly influenced by the nature and extent of available information. In financial markets, particularly foreign exchange (forex), traders operate independently with…

计算工程、金融与科学 · 计算机科学 2024-12-03 Patrick Naivasha , George Musumba , Patrick Gikunda , John Wandeto

Ensuring robust safety alignment is crucial for Large Language Models (LLMs), yet existing defenses often lag behind evolving adversarial attacks due to their \textbf{reliance on static, pre-collected data distributions}. In this paper, we…

人工智能 · 计算机科学 2026-02-09 Xiaoyu Wen , Zhida He , Han Qi , Ziyu Wan , Zhongtian Ma , Ying Wen , Tianhang Zheng , Xingcheng Xu , Chaochao Lu , Qiaosheng Zhang

We study the problem of repeated play in a zero-sum game in which the payoff matrix may change, in a possibly adversarial fashion, on each round; we call these Online Matrix Games. Finding the Nash Equilibrium (NE) of a two player zero-sum…

机器学习 · 计算机科学 2020-04-06 Adrian Rivera Cardoso , Jacob Abernethy , He Wang , Huan Xu

Repeated game has long been the touchstone model for agents' long-run relationships. Previous results suggest that it is particularly difficult for a repeated game player to exert an autocratic control on the payoffs since they are jointly…

计算机科学与博弈论 · 计算机科学 2018-07-19 Dong Hao , Kai Li , Tao Zhou

Proportionality is an attractive fairness concept that has been applied to a range of problems including the facility location problem, a classic problem in social choice. In our work, we propose a concept called Strong Proportionality,…

计算机科学与博弈论 · 计算机科学 2022-06-15 Haris Aziz , Alexander Lam , Mashbat Suzuki , Toby Walsh

We study the mechanism design problem of allocating a set of indivisible items without monetary transfers. Despite the vast literature on this very standard model, it still remains unclear how do truthful mechanisms look like. We focus on…

计算机科学与博弈论 · 计算机科学 2017-05-31 Georgios Amanatidis , Georgios Birmpas , George Christodoulou , Evangelos Markakis

According to the evolutionary game theory principle, a strategy representing a higher payoff can spread among competitors. But there are cases when a player consistently overestimates or underestimates her own payoff, which undermines…

物理与社会 · 物理学 2018-09-06 Attila Szolnoki , Xiaojie Chen

We consider control of heterogeneous players repeatedly playing an anti-coordination network game. In an anti-coordination game, each player has an incentive to differentiate its action from its neighbors. At each round of play, players…

系统与控制 · 计算机科学 2018-12-13 Ceyhun Eksin , Keith Paarporn

A knockout tournament is one of the most simple and popular forms of competition. Here, we are given a binary tournament tree where all leaves are labeled with seed position names. The players participating in the tournament are assigned to…

离散数学 · 计算机科学 2025-06-05 Klim Efremenko , Hendrik Molter , Meirav Zehavi