中文
相关论文

相关论文: Value Functions for Depth-Limited Solving in Zero-…

200 篇论文

Recent years have seen the application of deep reinforcement learning techniques to cooperative multi-agent systems, with great empirical success. However, given the lack of theoretical insight, it remains unclear what the employed neural…

多智能体系统 · 计算机科学 2024-12-20 Jacopo Castellini , Frans A. Oliehoek , Rahul Savani , Shimon Whiteson

Bayesian games model interactive decision-making where players have incomplete information -- e.g., regarding payoffs and private data on players' strategies and preferences -- and must actively reason and update their belief models (with…

计算机科学与博弈论 · 计算机科学 2024-05-24 Zuyuan Zhang , Mahdi Imani , Tian Lan

This paper extends Berge's maximum theorem for possibly noncompact action sets and unbounded cost functions to minimax problems and studies applications of these extensions to two-player zero-sum games with possibly noncompact action sets…

最优化与控制 · 数学 2017-09-14 Eugene A. Feinberg , Pavlo O. Kasyanov , Michael Z. Zgurovsky

In this paper, we consider the problem of optimization and learning for constrained and multi-objective Markov decision processes, for both discounted rewards and expected average rewards. We formulate the problems as zero-sum games where…

最优化与控制 · 数学 2021-03-05 Ather Gattami , Qinbo Bai , Vaneet Agarwal

This paper provides a decomposition technique for the purpose of simplifying the solution of certain zero-sum differential games. The games considered terminate when the state reaches a target, which can be expressed as the union of a…

最优化与控制 · 数学 2014-09-17 Adriano Festa , Richard Vinter

Decentralized team problems where players have asymmetric information about the state of the underlying stochastic system have been actively studied, but \emph{games} between such teams are less understood. We consider a general model of…

多智能体系统 · 计算机科学 2021-09-29 Dhruva Kartik , Ashutosh Nayyar , Urbashi Mitra

This paper studies two-player zero-sum repeated Bayesian games in which every player has a private type that is unknown to the other player, and the initial probability of the type of every player is publicly known. The types of players are…

计算机科学与博弈论 · 计算机科学 2017-11-08 Lichun Li , Cedric Langbort , Jeff Shamma

A class of nonzero-sum stochastic dynamic games with imperfect information structure is investigated. The game involves an arbitrary number of players, modeled as homogeneous Markov decision processes, aiming to find a sequential Nash…

最优化与控制 · 数学 2019-12-17 Jalal Arabneydi , Amir G. Aghdam

In this paper, we extend the Descent framework, which enables learning and planning in the context of two-player games with perfect information, to the framework of stochastic games. We propose two ways of doing this, the first way…

人工智能 · 计算机科学 2023-02-10 Quentin Cohen-Solal , Tristan Cazenave

A large body of research is currently investigating on the connection between machine learning and game theory. In this work, game theory notions are injected into a preference learning framework. Specifically, a preference learning problem…

机器学习 · 计算机科学 2018-12-20 Mirko Polato , Fabio Aiolli

Online learning algorithms that minimize regret provide strong guarantees in situations that involve repeatedly making decisions in an uncertain environment, e.g. a driver deciding what route to drive to work every day. While regret…

计算机科学与博弈论 · 计算机科学 2013-09-06 Jeremiah Blocki , Nicolas Christin , Anupam Datta , Arunesh Sinha

The information decomposition problem requires an additive decomposition of the mutual information between the input and target variables into nonnegative terms. The recently introduced solution to this problem, Information Attribution,…

信息论 · 计算机科学 2022-07-13 Tomáš Kroupa , Sara Vannucci , Tomáš Votroubek

This paper investigates a class of multi-player discrete games where each player aims to maximize its own utility function. Each player does not know the other players' action sets, their deployed actions or the structures of its own or the…

最优化与控制 · 数学 2017-12-05 Zhisheng Hu , Minghui Zhu , Ping Chen , Peng Liu

We study the performance of optimistic regret-minimization algorithms for both minimizing regret in, and computing Nash equilibria of, zero-sum extensive-form games. In order to apply these algorithms to extensive-form games, a…

计算机科学与博弈论 · 计算机科学 2019-10-29 Gabriele Farina , Christian Kroer , Tuomas Sandholm

We investigate the interrelation between graph searching games and games with imperfect information. As key consequence we obtain that parity games with bounded imperfect information can be solved in PTIME on graphs of bounded DAG-width…

计算机科学与博弈论 · 计算机科学 2015-03-19 Bernd Puchala , Roman Rabinovich

Agent-compiled knowledge bases provide persistent external knowledge for large language model (LLM) agents in open-ended, knowledge-intensive downstream tasks. Yet their quality is systematically limited by \emph{incompleteness},…

计算与语言 · 计算机科学 2026-05-12 Haoyu Huang , Jiaxin Bai , Shujie Liu , Yang Wei , Hong Ting Tsang , Yisen Gao , Zhongwei Xie , Yufei Li , Yangqiu Song

We study the iteration complexity of decentralized learning of approximate correlated equilibria in incomplete information games. On the negative side, we prove that in $\mathit{extensive}$-$\mathit{form}$ $\mathit{games}$, assuming…

计算机科学与博弈论 · 计算机科学 2024-06-05 Binghui Peng , Aviad Rubinstein

We introduce ZeroSumEval, a dynamic, competition-based, and evolving evaluation framework for Large Language Models (LLMs) that leverages competitive games. ZeroSumEval encompasses a diverse suite of games, including security challenges…

计算与语言 · 计算机科学 2025-04-18 Hisham A. Alyahya , Haidar Khan , Yazeed Alnumay , M Saiful Bari , Bülent Yener

In two-player zero-sum games, if both players minimize their average external regret, then the average of the strategy profiles converges to a Nash equilibrium. For n-player general-sum games, however, theoretical guarantees for regret…

计算机科学与博弈论 · 计算机科学 2013-05-02 Richard Gibson

We investigate optimal decision making under imperfect recall, that is, when an agent forgets information it once held before. An example is the absentminded driver game, as well as team games in which the members have limited communication…

计算机科学与博弈论 · 计算机科学 2024-06-25 Emanuel Tewolde , Brian Hu Zhang , Caspar Oesterheld , Manolis Zampetakis , Tuomas Sandholm , Paul W. Goldberg , Vincent Conitzer