中文
相关论文

相关论文: Anchoring Theory in Sequential Stackelberg Games

200 篇论文

We study the behavioral implications of Rationality and Common Strong Belief in Rationality (RCSBR) with contextual assumptions allowing players to entertain misaligned beliefs, i.e., players can hold beliefs concerning their opponents'…

理论经济学 · 经济学 2022-05-03 Pierfrancesco Guarino , Gabriel Ziegler

Reinforcement learning (RL) has recently proven effective at scaling chain-of-thought (CoT) reasoning in large language models for tasks with verifiable answers. However, extending RL-based thought training to more general non-verifiable…

We introduce and study a new model of interactive proofs: AM(k), or Arthur-Merlin with k non-communicating Merlins. Unlike with the better-known MIP, here the assumption is that each Merlin receives an independent random challenge from…

计算复杂性 · 计算机科学 2014-01-28 Scott Aaronson , Russell Impagliazzo , Dana Moshkovitz

Asymmetric information stochastic games (AISGs) arise in many complex socio-technical systems, such as cyber-physical systems and IT infrastructures. Existing computational methods for AISGs are primarily offline and can not adapt to…

计算机科学与博弈论 · 计算机科学 2024-08-20 Tao Li , Kim Hammar , Rolf Stadler , Quanyan Zhu

We consider a one-round two-player network pricing game, the Stackelberg Minimum Spanning Tree game or StackMST. The game is played on a graph (representing a network), whose edges are colored either red or blue, and where the red edges…

计算机科学与博弈论 · 计算机科学 2011-03-07 Jean Cardinal , Erik D. Demaine , Samuel Fiorini , Gwenaël Joret , Stefan Langerman , Ilan Newman , Oren Weimann

The present survey aims at presenting the current machine learning techniques employed in security games domains. Specifically, we focused on papers and works developed by the Teamcore of University of Southern California, which deepened…

计算机科学与博弈论 · 计算机科学 2016-09-30 Giuseppe De Nittis , Francesco Trovò

Equilibrium refinements are important in extensive-form (i.e., tree-form) games, where they amend weaknesses of the Nash equilibrium concept by requiring sequential rationality and other beneficial properties. One of the most attractive…

计算机科学与博弈论 · 计算机科学 2018-11-12 Alberto Marchesi , Gabriele Farina , Christian Kroer , Nicola Gatti , Tuomas Sandholm

The paper [Ras15a] introduced distribution-valued games. This game-theoretic model uses probability distributions as payoffs for games in order to express uncertainty about the payoffs. The player's preferences for different payoffs are…

最优化与控制 · 数学 2021-03-26 Vincent Bürgin

We consider game-theoretically secure distributed protocols for coalition games that approximate the Shapley value with small multiplicative error. Since all known existing approximation algorithms for the Shapley value are randomized, it…

计算机科学与博弈论 · 计算机科学 2024-12-30 T-H. Hubert Chan , Qipeng Kuang , Quan Xue

We study payoff manipulation in repeated multi-objective Stackelberg games, where a leader may strategically influence a follower's deterministic best response, e.g., by offering a share of their own payoff. We assume that the follower's…

计算机科学与博弈论 · 计算机科学 2025-08-27 Phurinut Srisawad , Juergen Branke , Long Tran-Thanh

A growing body of work in game theory extends the traditional Stackelberg game to settings with one leader and multiple followers who play a Nash equilibrium. Standard approaches for computing equilibria in these games reformulate the…

计算机科学与博弈论 · 计算机科学 2021-12-07 Kai Wang , Lily Xu , Andrew Perrault , Michael K. Reiter , Milind Tambe

In recent years, Signal Temporal Logic (STL) has gained traction as a practical and expressive means of encoding control objectives for robotic and cyber-physical systems. The state-of-the-art in STL trajectory synthesis is to formulate the…

机器人学 · 计算机科学 2019-05-09 Vince Kurtz , Hai Lin

TheMinority Game (MG) has become a paradigm to probe complex social and economical phenomena where adaptive agents compete for a limited resource, and it finds applications in statistical and nonlinear physics as well. In the traditional MG…

适应与自组织系统 · 物理学 2012-04-16 Zi-Gang Huang , Ji-Qiang Zhang , Jia-Qi Dong , Liang Huang , Ying-Cheng Lai

The takeoff point for this paper is the voluminous body of literature addressing recursive betting games with expected logarithmic growth of wealth being the performance criterion. Whereas almost all existing papers involve use of linear…

最优化与控制 · 数学 2024-01-17 Anton V. Proskurnikov , B. Ross Barmish

Chain-of-thought (CoT) reasoning with self-consistency improves performance by aggregating multiple sampled reasoning paths. In this setting, correctness is no longer tied to a single reasoning trace but to the aggregation rule over a pool…

机器学习 · 统计学 2026-05-15 Yu Gu , Zijun Yu , Vahid Partovi Nia , Masoud Asgharian

This paper addresses a Stackelberg stochastic linear-quadratic (LQ) differential game under closed-loop information, a problem inherently time-inconsistent. Existing approaches rely on solving two coupled Hamilton-Jacobi-Bellman (HJB)…

最优化与控制 · 数学 2026-04-27 Qi Lü , Bowen Ma , Hanxiao Wang

Driven by recent successes in two-player, zero-sum game solving and playing, artificial intelligence work on games has increasingly focused on algorithms that produce equilibrium-based strategies. However, this approach has been less…

计算机科学与博弈论 · 计算机科学 2022-06-24 Dustin Morrill , Ryan D'Orazio , Reca Sarfati , Marc Lanctot , James R. Wright , Amy Greenwald , Michael Bowling

We present a robust framework with computational algorithms to support decision makers in sequential games. Our framework includes methods to solve games with complete information, assess the robustness of such solutions and, finally,…

统计计算 · 统计学 2024-02-22 Tahir Ekin , Roi Naveiro , Alberto Torres-Barrán , David Ríos-Insua

We propose a sequential optimizing betting strategy in the multi-dimensional bounded forecasting game in the framework of game-theoretic probability of Shafer and Vovk (2001). By studying the asymptotic behavior of its capital process, we…

概率论 · 数学 2011-02-16 Masayuki Kumon , Akimichi Takemura , Kei Takeuchi

We consider multi-armed bandit problems in social groups wherein each individual has bounded memory and shares the common goal of learning the best arm/option. We say an individual learns the best option if eventually (as $t\to \infty$) it…

分布式、并行与集群计算 · 计算机科学 2018-12-27 Lili Su , Martin Zubeldia , Nancy Lynch