中文
相关论文

相关论文: On the complexity of heterogeneous multidimensiona…

200 篇论文

We consider the problem of learning to play a repeated multi-agent game with an unknown reward function. Single player online learning algorithms attain strong regret bounds when provided with full information feedback, which unfortunately…

机器学习 · 计算机科学 2019-10-29 Pier Giuseppe Sessa , Ilija Bogunovic , Maryam Kamgarpour , Andreas Krause

In many multiagent environments, a designer has some, but limited control over the game being played. In this paper, we formalize this by considering incompletely specified games, in which some entries of the payoff matrices can be chosen…

计算机科学与博弈论 · 计算机科学 2021-04-30 Markus Brill , Rupert Freeman , Vincent Conitzer

Extensive-form games with imperfect recall are an important game-theoretic model that allows a compact representation of strategies in dynamic strategic interactions. Practical use of imperfect recall games is limited due to negative…

计算机科学与博弈论 · 计算机科学 2017-05-25 Branislav Bosansky , Jiri Cermak , Karel Horak , Michal Pechoucek

This paper studies two-player zero-sum games played on graphs and makes contributions toward the following question: given an objective, how much memory is required to play optimally for that objective? We study regular objectives, where…

计算机科学与博弈论 · 计算机科学 2023-09-19 Patricia Bouyer , Nathanaël Fijalkow , Mickael Randour , Pierre Vandenhove

Robust Markov decision processes (RMDPs) extend standard Markov decision processes (MDPs) to account for uncertainty in the transition probabilities. RMDPs have an uncertainty set that defines a set of possible transition functions, each of…

计算机科学中的逻辑 · 计算机科学 2026-04-30 Marnix Suilen , Guillermo A. Pérez

Large Language Models (LLMs) define probability measures on text. By considering the implicit knowledge question of what it means for an LLM to know such a measure and what it entails algorithmically, we are naturally led to formulate a…

人工智能 · 计算机科学 2025-06-24 Clément Hongler , Andrew Emil

We introduce quantitative reductions, a novel technique for structuring the space of quantitative games and solving them that does not rely on a reduction to qualitative games. We show that such reductions exhibit the same desirable…

计算机科学与博弈论 · 计算机科学 2018-09-12 Alexander Weinert

We study synthesis problems with constraints in partially observable Markov decision processes (POMDPs), where the objective is to compute a strategy for an agent that is guaranteed to satisfy certain safety and performance specifications.…

We consider two-player games played over finite state spaces for an infinite number of rounds. At each state, the players simultaneously choose moves; the moves determine a successor state. It is often advantageous for players to choose…

计算机科学中的逻辑 · 计算机科学 2015-07-01 Luca de Alfaro , Rupak Majumdar , Vishwanath Raman , Mariëlle Stoelinga

This paper investigates the problem of computing the equilibrium of competitive games, which is often modeled as a constrained saddle-point optimization problem with probability simplex constraints. Despite recent efforts in understanding…

最优化与控制 · 数学 2023-01-23 Shicong Cen , Yuting Wei , Yuejie Chi

This paper studies two-player zero-sum stochastic Bayesian games where each player has its own dynamic state that is unknown to the other player. Using typical techniques, we provide the recursive formulas and sufficient statistics in both…

计算机科学与博弈论 · 计算机科学 2021-05-05 Nabiha Nasir Orpa , Lichun Li

The extensive-form game has been studied considerably in recent years. It can represent games with multiple decision points and incomplete information, and hence it is helpful in formulating games with uncertain inputs, such as poker. We…

计算机科学与博弈论 · 计算机科学 2023-03-21 Keigo Habara , Ellen Hidemi Fukuda , Nobuo Yamashita

We extend the quantitative synthesis framework by going beyond the worst-case. On the one hand, classical analysis of two-player games involves an adversary (modeling the environment of the system) which is purely antagonistic and asks for…

计算机科学与博弈论 · 计算机科学 2015-11-02 Véronique Bruyère , Emmanuel Filiot , Mickael Randour , Jean-François Raskin

An approximation algorithm for a constraint satisfaction problem is called robust if it outputs an assignment satisfying a $(1 - f(\epsilon))$-fraction of the constraints on any $(1-\epsilon)$-satisfiable instance, where the loss function…

数据结构与算法 · 计算机科学 2022-11-09 Antoine Méot , Arnaud de Mesmay , Moritz Mühlenthaler , Alantha Newman

We study best-response type learning dynamics for zero-sum polymatrix games under two information settings. The two settings are distinguished by the type of information that each player has about the game and their opponents' strategy. The…

最优化与控制 · 数学 2025-08-13 Fathima Zarin Faizal , Asuman Ozdaglar , Martin J. Wainwright

Boolean games are an expressive and natural formalism through which to investigate problems of strategic interaction in multiagent systems. Although they have been widely studied, almost all previous work on Nash equilibria in Boolean games…

计算机科学与博弈论 · 计算机科学 2013-12-17 Egor Ianovski , Luke Ong

Cheung and Piliouras (2020) recently showed that two variants of the Multiplicative Weights Update method - OMWU and MWU - display opposite convergence properties depending on whether the game is zero-sum or cooperative. Inspired by this…

计算机科学与博弈论 · 计算机科学 2022-06-14 Nelson Vadori , Rahul Savani , Thomas Spooner , Sumitra Ganesh

Games on recursive game graphs can be used to reason about the control flow of sequential programs with recursion. In games over recursive game graphs, the most natural notion of strategy is the modular strategy, i.e., a strategy that is…

计算机科学中的逻辑 · 计算机科学 2014-08-27 Ilaria De Crescenzo , Salvatore La Torre , Yaron Velner

A central task of artificial intelligence is the design of artificial agents that act towards specified goals in partially observed environments. Since such environments frequently include interaction over time with other agents with their…

计算机科学与博弈论 · 计算机科学 2012-05-14 Miroslav Dudik , Geoffrey Gordon

We consider two-player games played on finite colored graphs where the goal is the construction of an infinite path with one of the following frequency-related properties: (i) all colors occur with the same asymptotic frequency, (ii) there…

计算机科学与博弈论 · 计算机科学 2010-06-29 Alessandro Bianco , Marco Faella , Fabio Mogavero , Aniello Murano
‹ 上一页 1 8 9 10 下一页 ›