中文
相关论文

相关论文: A Stochastic Linear-Quadratic Leader-Follower Diff…

200 篇论文

Resource competition problems are often modeled using Colonel Blotto games, where players take simultaneous actions. However, many real-world scenarios involve sequential decision-making rather than simultaneous moves. To model these…

计算机科学与博弈论 · 计算机科学 2025-05-13 Yan Liu , Bonan Ni , Weiran Shen , Zihe Wang , Jie Zhang

A Stackelberg game is played between a leader and a follower. The leader first chooses an action, then the follower plays his best response. The goal of the leader is to pick the action that will maximize his payoff given the follower's…

数据结构与算法 · 计算机科学 2015-11-19 Aaron Roth , Jonathan Ullman , Zhiwei Steven Wu

Our purpose of this paper is to study stochastic control problem for systems driven by mean-field stochastic differential equations with elephant memory, in the sense that the system (like the elephants) never forgets its history. We study…

最优化与控制 · 数学 2019-06-24 Nacira Agram , Bernt Øksendal

Information uncertainty is one of the major challenges facing applications of game theory. In the context of Stackelberg games, various approaches have been proposed to deal with the leader's incomplete knowledge about the follower's…

计算机科学与博弈论 · 计算机科学 2019-05-21 Jiarui Gan , Haifeng Xu , Qingyu Guo , Long Tran-Thanh , Zinovi Rabinovich , Michael Wooldridge

Guided cooperation allows intelligent agents with heterogeneous capabilities to work together by following a leader-follower type of interaction. However, the associated control problem becomes challenging when the leader agent does not…

系统与控制 · 电气工程与系统科学 2024-02-01 Yuhan Zhao , Quanyan Zhu

We introduce a stochastic principal-agent model. A principal and an agent interact in a stochastic environment, each privy to observations about the state not available to the other. The principal has the power of commitment, both to elicit…

计算机科学与博弈论 · 计算机科学 2024-09-13 Jiarui Gan , Rupak Majumdar , Debmalya Mandal , Goran Radanovic

We introduce a two-player model of reinforcement learning with memory. Past actions of an iterated game are stored in a memory and used to determine player's next action. To examine the behaviour of the model some approximate methods are…

统计力学 · 物理学 2009-11-13 Adam Lipowski , Krzysztof Gontarek , Marcel Ausloos

In a multi-follower Bayesian Stackelberg game, a leader plays a mixed strategy over $L$ actions to which $n\ge 1$ followers, each having one of $K$ possible private types, best respond. The leader's optimal strategy depends on the…

计算机科学与博弈论 · 计算机科学 2026-03-03 Gerson Personnat , Tao Lin , Safwan Hossain , David C. Parkes

We introduce an original way to estimate the memory parameter of the elephant random walk, a fascinating discrete time random walk on integers having a complete memory of its entire history. Our estimator is nothing more than a…

概率论 · 数学 2021-12-21 Bernard Bercu , Lucile Laulin

Elephant random walk, introduced to study the effect of memory on random walks, is a novel type of walk that incorporates the information of one randomly chosen past step to determine the future step. However, memory of a process can be…

概率论 · 数学 2025-09-15 Krishanu Maulik , Parthanil Roy , Tamojit Sadhukhan

Dynamic Stackelberg games are a broad class of two-player games in which the leader acts first, and the follower chooses a response strategy to the leader's strategy. Unfortunately, only stylized Stackelberg games are explicitly solvable…

最优化与控制 · 数学 2024-11-15 Guillermo Alvarez , Ibrahim Ekren , Anastasis Kratsios , Xuwei Yang

We explore the impact of long-range memory on the properties of a family of quantum walks in a one-dimensional lattice and discrete time, which can be understood as the quantum version of the classical "Elephant Random Walk" non-Markovian…

量子物理 · 物理学 2018-06-20 Giuseppe Di Molfetta , Diogo O. Soares-Pinto , Silvio M. Duarte Queiros

We consider a general time-inconsistent stochastic linear-quadratic differential game. The time-inconsistency arises from the presence of quadratic terms of the expected state as well as state-dependent term in the objective functionals. We…

数理金融 · 定量金融 2024-05-15 Qinglong Zhou , Gaofeng Zong

We consider a repeated Stackelberg game setup where the leader faces a sequence of followers of unknown types and must learn what commitments to make. While previous works have considered followers that best respond to the commitment…

计算机科学与博弈论 · 计算机科学 2024-12-10 Vijeth Hebbar , Cédric Langbort

This paper studies a new class of dynamic optimization problems of large-population (LP) system which consists of a large number of negligible and coupled agents. The most significant feature in our setup is the dynamics of individual…

最优化与控制 · 数学 2014-03-18 Jianhui Huang , Shujun Wang , Hua Xiao

This paper is concerned with a class of linear-quadratic stochastic large-population problems with partial information, where the individual agent only has access to a noisy observation process related to the state. The dynamics of each…

最优化与控制 · 数学 2024-08-20 Min Li , Na Li , Zhen Wu

We study Stackelberg equilibria in finitely repeated games, where the leader commits to a strategy that picks actions in each round and can be adaptive to the history of play (i.e. they commit to an algorithm). In particular, we study…

计算机科学与博弈论 · 计算机科学 2024-03-08 Natalie Collina , Eshwar Ram Arunachaleswaran , Michael Kearns

The elephant random walk (ERW) is a microscopic, one-dimensional, discrete-time, non-Markovian random walk, which can lead to anomalous diffusion due to memory effects. In this study, I propose a multi-dimensional generalization in which…

统计力学 · 物理学 2019-12-02 Vitor M. Marquioni

We study payoff manipulation in repeated multi-objective Stackelberg games, where a leader may strategically influence a follower's deterministic best response, e.g., by offering a share of their own payoff. We assume that the follower's…

计算机科学与博弈论 · 计算机科学 2025-08-27 Phurinut Srisawad , Juergen Branke , Long Tran-Thanh

We study incentive designs for a class of stochastic Stackelberg games with one leader and a large number of (finite as well as infinite population of) followers. We investigate whether the leader can craft a strategy under a dynamic…

计算机科学与博弈论 · 计算机科学 2024-02-13 Sina Sanjari , Subhonmesh Bose , Tamer Başar