中文
相关论文

相关论文: Promises Made, Promises Kept: Safe Pareto Improvem…

200 篇论文

The conditional commitment abilities of mutually transparent computer agents have been studied in previous work on commitment games and program equilibrium. This literature has shown how these abilities can help resolve Prisoner's Dilemmas…

计算机科学与博弈论 · 计算机科学 2022-12-06 Anthony DiGiovanni , Jesse Clifton

Our ability to know when to trust the decisions made by machine learning systems has not kept up with the staggering improvements in their performance, limiting their applicability in high-stakes domains. We introduce Prover-Verifier Games…

机器学习 · 计算机科学 2021-08-30 Cem Anil , Guodong Zhang , Yuhuai Wu , Roger Grosse

Many problems in compositional synthesis and verification of multi-agent systems -- such as rational verification and assume-guarantee verification in probabilistic systems -- reduce to reasoning about two-player multi-objective stochastic…

计算机科学与博弈论 · 计算机科学 2026-02-16 Moritz Graf , Anthony Lin , Rupak Majumdar

Every interaction of a living organism with its environment involves the placement of a bet. Armed with partial knowledge about a stochastic world, the organism must decide its next step or near-term strategy, an act that implicitly or…

种群与进化 · 定量生物学 2023-05-30 Philipp Fleig , Vijay Balasubramanian

Symmetric strategy improvement is an algorithm introduced by Schewe et al. (ICALP 2015) that can be used to solve two-player games on directed graphs such as parity games and mean payoff games. In contrast to the usual well-known strategy…

计算机科学与博弈论 · 计算机科学 2023-09-06 Tom van Dijk , Georg Loho , Matthew Maat

We introduce games with probabilistic uncertainty, a natural model for controller synthesis in which the controller observes the state of the system through imprecise sensors that provide correct information about the current state with a…

计算机科学与博弈论 · 计算机科学 2012-07-03 Krishnendu Chatterjee , Martin Chmelik , Rupak Majumdar

We study discrete preference games in heterogeneous social networks. These games model the interplay between a player's private belief and his/her publicly stated opinion (which could be different from the player's belief) as a strategic…

计算机科学与博弈论 · 计算机科学 2016-03-10 Vincenzo Auletta , Ioannis Caragiannis , Diodato Ferraioli , Clemente Galdi , Giuseppe Persiano

Commitment devices are powerful tools that can influence and incentivise certain behaviours by linking them to rewards or punishments. These devices are particularly useful in decision-making, as they can steer individuals towards specific…

计算机科学与博弈论 · 计算机科学 2023-12-11 Maria Alejandra Ramirez , Yoav Kolumbus , Rosemarie Nagel , David Wolpert , Jürgen Jost

This paper introduces an explicit algorithm for computing perfect public equilibrium (PPE) payoffs in repeated games with imperfect public monitoring, public randomization, and discounting. The method adapts the established framework by…

理论经济学 · 经济学 2024-11-05 Jasmina Karabegovic

Designing hierarchical reinforcement learning algorithms that exhibit safe behaviour is not only vital for practical applications but also, facilitates a better understanding of an agent's decisions. We tackle this problem in the options…

人工智能 · 计算机科学 2021-07-01 Arushi Jain , Khimya Khetarpal , Doina Precup

In this paper, we study the problem of learning the exact structure of continuous-action games with non-parametric utility functions. We propose an $\ell_1$ regularized method which encourages sparsity of the coefficients of the Fourier…

计算机科学与博弈论 · 计算机科学 2022-09-16 Adarsh Barik , Jean Honorio

I prove that it is irrational for agents with even slightly private preferences to condition their strategy on private information that is payoff-irrelevant to them, contrary to powerful techniques for analyzing communication and repeated…

理论经济学 · 经济学 2026-05-29 Alistair Barton

We consider finite-horizon and infinite-horizon versions of a dynamic game with $N$ selfish players who observe their types privately and take actions that are publicly observed. Players' types evolve as conditionally independent Markov…

最优化与控制 · 数学 2018-03-20 Deepanshu Vasal , Abhinav Sinha , Achilleas Anastasopoulos

We study the variant of the stable marriage problem in which the preferences of the agents are allowed to include indifferences. We present a mechanism for producing Pareto-stable matchings in stable marriage markets with indifferences that…

计算机科学与博弈论 · 计算机科学 2017-10-13 Nevzat Onur Domaniç , Chi-Kit Lam , C. Gregory Plaxton

We study a sequence of independent one-shot non-cooperative games where agents play equilibria determined by a tunable mechanism. Observing only equilibrium decisions, without parametric or distributional knowledge of utilities, we aim to…

计算机科学与博弈论 · 计算机科学 2025-11-10 Luke Snow , Vikram Krishnamurthy

We study finite normal-form games in which payoffs are subject to random perturbations and players face uncertainty about how these shocks co-move across actions, an ambiguity that naturally arises when only realized (not counterfactual)…

理论经济学 · 经济学 2026-02-12 Yu Gui , Bahar Taşkesen

Behavioral experiments on the Ultimatum Game have shown that we human beings have remarkable preference in fair play, contradicting the predictions by the game theory. Most of the existing models seeking for explanations, however, strictly…

种群与进化 · 定量生物学 2022-12-07 Guozhong Zheng , Jiqiang Zhang , Rizhou Liang , Lin Ma , Li Chen

In a typical school choice application, the students have strict preferences over the schools while the schools have coarse priorities over the students based on their distance and their enrolled siblings. The outcome of a centralized…

计算机科学与博弈论 · 计算机科学 2026-02-12 Haris Aziz , Péter Biró , Gergely Csáji , Tom Demeulemeester

While discounted payoff games and classic games that reduce to them, like parity and mean-payoff games, are symmetric, their solutions are not. We have taken a fresh view on the constraints that optimal solutions need to satisfy, and…

数据结构与算法 · 计算机科学 2023-10-03 Daniele Dell'Erba , Arthur Dumas , Sven Schewe

This paper presents a technique for approximating, up to any precision, the set of subgame-perfect equilibria (SPE) in discounted repeated games. The process starts with a single hypercube approximation of the set of SPE. Then the initial…

计算机科学与博弈论 · 计算机科学 2010-02-10 Andriy Burkov , Brahim Chaib-draa