English
Related papers

Related papers: Markovian Persuasion with Two States

200 papers

We present a new behavioural distance over the state space of a Markov decision process, and demonstrate the use of this distance as an effective means of shaping the learnt representations of deep reinforcement learning agents. While…

Machine Learning · Computer Science 2022-01-25 Pablo Samuel Castro , Tyler Kastner , Prakash Panangaden , Mark Rowland

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

Machine Learning · Computer Science 2025-12-24 Peter N. Loxley

The general picture of game theoretic modeling dealt with here is characterized by a set of big players, also referred to as principals or major agents, acting on the background of large pools of small players, the impact of the behavior of…

Optimization and Control · Mathematics 2019-11-12 Vassili N. Kolokoltsov , Oleg A. Malafeyev

Starting from the Avellaneda-Stoikov framework, we consider a market maker who wants to optimally set bid/ask quotes over a finite time horizon, to maximize her expected utility. The intensities of the orders she receives depend not only on…

Trading and Market Microstructure · Quantitative Finance 2020-06-29 Diego Zabaljauregui , Luciano Campi

The optimal control for mobile agents is an important and challenging issue. Recent work shows that using randomized mechanism in agents' control can make the state unpredictable, and thus improve the security of agents. However, the…

Systems and Control · Electrical Eng. & Systems 2022-09-05 Chendi Qu , Jianping He , Jialun Li

This paper presents a novel state representation for reward-free Markov decision processes. The idea is to learn, in a self-supervised manner, an embedding space where distances between pairs of embedded states correspond to the minimum…

Machine Learning · Computer Science 2022-05-05 Lorenzo Steccanella , Anders Jonsson

We present a model that explores the influence of persuasion in a population of agents with positive and negative opinion orientations. The opinion of each agent is represented by an integer number $k$ that expresses its level of agreement…

Physics and Society · Physics 2014-05-30 C. E. La Rocca , L. A. Braunstein , F. Vazquez

Analytical approaches in models of opinion formation have been extensively studied either for an opinion represented as a discrete or a continuous variable. In this paper, we analyze a model which combines both approaches. The state of an…

Physics and Society · Physics 2023-02-01 Lucía Pedraza , Juan Pablo Pinasco , Viktoriya Semeshenko , Pablo Balenzuela

In this work we study the optimal execution problem with multiplicative price impact in algorithm trading, when an agent holds an initial position of shares of a financial asset. The inter-selling-decision times are modelled by the arrival…

Mathematical Finance · Quantitative Finance 2018-05-04 Daniel Hernández-Hernández , Harold A. Moreno-Franco , José Luis Pérez

In multi-agent reinforcement learning, the problem of learning to act is particularly difficult because the policies of co-players may be heavily conditioned on information only observed by them. On the other hand, humans readily form…

Machine Learning · Computer Science 2021-02-05 Pol Moreno , Edward Hughes , Kevin R. McKee , Bernardo Avila Pires , Théophane Weber

In a zero-sum stochastic game with signals, at each stage, two adversary players take decisions and receive a stage payoff determined by these decisions and a variable called state. The state follows a Markov chain, that is controlled by…

Optimization and Control · Mathematics 2021-12-02 Bruno Ziliotto

Many decision problems in economics, information technology, and industry can be transformed to an optimal stopping of adapted random vectors with some utility function over the set of Markov times with respect to filtration build by the…

Optimization and Control · Mathematics 2020-11-04 Krzysztof Szajowski

We study strategic information transmission in a hierarchical setting where information gets transmitted through a chain of agents up to a decision maker whose action is of importance to every agent. This situation could arise whenever an…

Theoretical Economics · Economics 2022-07-14 Majid Mahzoon

General purpose intelligent learning agents cycle through (complex,non-MDP) sequences of observations, actions, and rewards. On the other hand, reinforcement learning is well-developed for small finite state Markov Decision Processes…

Artificial Intelligence · Computer Science 2009-12-30 Marcus Hutter

When humans are given a policy to execute, there can be policy execution errors and deviations in policy if there is uncertainty in identifying a state. This can happen due to the human agent's cognitive limitations and/or perceptual…

Artificial Intelligence · Computer Science 2022-03-07 Sriram Gopalakrishnan , Mudit Verma , Subbarao Kambhampati

This paper looks at predictability problems, i.e., wherein an agent must choose its strategy in order to optimize the predictions that an external observer could make. We address these problems while taking into account uncertainties on the…

Artificial Intelligence · Computer Science 2024-10-08 Salomé Lepers , Sophie Lemonnier , Vincent Thomas , Olivier Buffet

We study a model of binary decision making when a certain population of agents is initially seeded with two different opinions, `$+$' and `$-$', with fractions $p_1$ and $p_2$ respectively, $p_1+p_2=1$. Individuals can reverse their initial…

Physics and Society · Physics 2018-10-16 Amitava Datta

We develop a model-free approach to optimally control stochastic, Markovian systems subject to a reach-avoid constraint. Specifically, the state trajectory must remain within a safe set while reaching a target set within a finite time…

Optimization and Control · Mathematics 2025-09-30 Tingting Ni , Maryam Kamgarpour

This paper studies a multi-player, general-sum stochastic game characterized by a dual-stage temporal structure per period. The agents face uncertainty regarding the time-evolving state that is realized at the beginning of each period.…

Computer Science and Game Theory · Computer Science 2023-10-09 Tao Zhang , Quanyan Zhu

We present an approach to reduce the communication required between agents in a Multi-Agent learning system by exploiting the inherent robustness of the underlying Markov Decision Process. We compute so-called robustness surrogate functions…

Multiagent Systems · Computer Science 2022-09-08 Daniel Jarne Ornia , Manuel Mazo