English
Related papers

Related papers: Anchoring Theory in Sequential Stackelberg Games

200 papers

We study the behavioral implications of Rationality and Common Strong Belief in Rationality (RCSBR) with contextual assumptions allowing players to entertain misaligned beliefs, i.e., players can hold beliefs concerning their opponents'…

Theoretical Economics · Economics 2022-05-03 Pierfrancesco Guarino , Gabriel Ziegler

Reinforcement learning (RL) has recently proven effective at scaling chain-of-thought (CoT) reasoning in large language models for tasks with verifiable answers. However, extending RL-based thought training to more general non-verifiable…

We introduce and study a new model of interactive proofs: AM(k), or Arthur-Merlin with k non-communicating Merlins. Unlike with the better-known MIP, here the assumption is that each Merlin receives an independent random challenge from…

Computational Complexity · Computer Science 2014-01-28 Scott Aaronson , Russell Impagliazzo , Dana Moshkovitz

Asymmetric information stochastic games (AISGs) arise in many complex socio-technical systems, such as cyber-physical systems and IT infrastructures. Existing computational methods for AISGs are primarily offline and can not adapt to…

Computer Science and Game Theory · Computer Science 2024-08-20 Tao Li , Kim Hammar , Rolf Stadler , Quanyan Zhu

We consider a one-round two-player network pricing game, the Stackelberg Minimum Spanning Tree game or StackMST. The game is played on a graph (representing a network), whose edges are colored either red or blue, and where the red edges…

Computer Science and Game Theory · Computer Science 2011-03-07 Jean Cardinal , Erik D. Demaine , Samuel Fiorini , Gwenaël Joret , Stefan Langerman , Ilan Newman , Oren Weimann

The present survey aims at presenting the current machine learning techniques employed in security games domains. Specifically, we focused on papers and works developed by the Teamcore of University of Southern California, which deepened…

Computer Science and Game Theory · Computer Science 2016-09-30 Giuseppe De Nittis , Francesco Trovò

Equilibrium refinements are important in extensive-form (i.e., tree-form) games, where they amend weaknesses of the Nash equilibrium concept by requiring sequential rationality and other beneficial properties. One of the most attractive…

Computer Science and Game Theory · Computer Science 2018-11-12 Alberto Marchesi , Gabriele Farina , Christian Kroer , Nicola Gatti , Tuomas Sandholm

The paper [Ras15a] introduced distribution-valued games. This game-theoretic model uses probability distributions as payoffs for games in order to express uncertainty about the payoffs. The player's preferences for different payoffs are…

Optimization and Control · Mathematics 2021-03-26 Vincent Bürgin

We consider game-theoretically secure distributed protocols for coalition games that approximate the Shapley value with small multiplicative error. Since all known existing approximation algorithms for the Shapley value are randomized, it…

Computer Science and Game Theory · Computer Science 2024-12-30 T-H. Hubert Chan , Qipeng Kuang , Quan Xue

We study payoff manipulation in repeated multi-objective Stackelberg games, where a leader may strategically influence a follower's deterministic best response, e.g., by offering a share of their own payoff. We assume that the follower's…

Computer Science and Game Theory · Computer Science 2025-08-27 Phurinut Srisawad , Juergen Branke , Long Tran-Thanh

A growing body of work in game theory extends the traditional Stackelberg game to settings with one leader and multiple followers who play a Nash equilibrium. Standard approaches for computing equilibria in these games reformulate the…

Computer Science and Game Theory · Computer Science 2021-12-07 Kai Wang , Lily Xu , Andrew Perrault , Michael K. Reiter , Milind Tambe

In recent years, Signal Temporal Logic (STL) has gained traction as a practical and expressive means of encoding control objectives for robotic and cyber-physical systems. The state-of-the-art in STL trajectory synthesis is to formulate the…

Robotics · Computer Science 2019-05-09 Vince Kurtz , Hai Lin

TheMinority Game (MG) has become a paradigm to probe complex social and economical phenomena where adaptive agents compete for a limited resource, and it finds applications in statistical and nonlinear physics as well. In the traditional MG…

Adaptation and Self-Organizing Systems · Physics 2012-04-16 Zi-Gang Huang , Ji-Qiang Zhang , Jia-Qi Dong , Liang Huang , Ying-Cheng Lai

The takeoff point for this paper is the voluminous body of literature addressing recursive betting games with expected logarithmic growth of wealth being the performance criterion. Whereas almost all existing papers involve use of linear…

Optimization and Control · Mathematics 2024-01-17 Anton V. Proskurnikov , B. Ross Barmish

Chain-of-thought (CoT) reasoning with self-consistency improves performance by aggregating multiple sampled reasoning paths. In this setting, correctness is no longer tied to a single reasoning trace but to the aggregation rule over a pool…

Machine Learning · Statistics 2026-05-15 Yu Gu , Zijun Yu , Vahid Partovi Nia , Masoud Asgharian

This paper addresses a Stackelberg stochastic linear-quadratic (LQ) differential game under closed-loop information, a problem inherently time-inconsistent. Existing approaches rely on solving two coupled Hamilton-Jacobi-Bellman (HJB)…

Optimization and Control · Mathematics 2026-04-27 Qi Lü , Bowen Ma , Hanxiao Wang

Driven by recent successes in two-player, zero-sum game solving and playing, artificial intelligence work on games has increasingly focused on algorithms that produce equilibrium-based strategies. However, this approach has been less…

Computer Science and Game Theory · Computer Science 2022-06-24 Dustin Morrill , Ryan D'Orazio , Reca Sarfati , Marc Lanctot , James R. Wright , Amy Greenwald , Michael Bowling

We present a robust framework with computational algorithms to support decision makers in sequential games. Our framework includes methods to solve games with complete information, assess the robustness of such solutions and, finally,…

Computation · Statistics 2024-02-22 Tahir Ekin , Roi Naveiro , Alberto Torres-Barrán , David Ríos-Insua

We propose a sequential optimizing betting strategy in the multi-dimensional bounded forecasting game in the framework of game-theoretic probability of Shafer and Vovk (2001). By studying the asymptotic behavior of its capital process, we…

Probability · Mathematics 2011-02-16 Masayuki Kumon , Akimichi Takemura , Kei Takeuchi

We consider multi-armed bandit problems in social groups wherein each individual has bounded memory and shares the common goal of learning the best arm/option. We say an individual learns the best option if eventually (as $t\to \infty$) it…

Distributed, Parallel, and Cluster Computing · Computer Science 2018-12-27 Lili Su , Martin Zubeldia , Nancy Lynch