English
Related papers

Related papers: Entropic Risk for Turn-Based Stochastic Games

200 papers

The game-theoretic risk management framework put forth in the precursor work "Towards a Theory of Games with Payoffs that are Probability-Distributions" (arXiv:1506.07368 [q-fin.EC]) is herein extended by algorithmic details on how to…

General Economics · Economics 2020-04-10 Stefan Rass

We consider an investment problem in which an investor performs capital injections to increase the liquidity of a firm for it to maximise profit from market operations. Each time the investor performs an injection, the investor incurs a…

Optimization and Control · Mathematics 2019-10-04 David Mguni

Remote estimation is a crucial element of real time monitoring of a stochastic process. While most of the existing works have concentrated on obtaining optimal sampling strategies, motivated by malicious attacks on cyber-physical systems,…

Information Theory · Computer Science 2024-12-03 Atahan Dokme , Raj Kiriti Velicheti , Melih Bastopcu , Tamer Başar

We study a class of two-player zero-sum stochastic games known as \textit{blind stochastic games}, where players neither observe the state nor receive any information about it during the game. A central concept for analyzing long-duration…

Optimization and Control · Mathematics 2025-11-24 Krishnendu Chatterjee , David Lurie , Raimundo Saona , Bruno Ziliotto

This paper investigates methods for estimating the optimal stochastic control policy for a Markov Decision Process with unknown transition dynamics and an unknown reward function. This form of model-free reinforcement learning comprises…

Machine Learning · Computer Science 2019-12-06 Brandon Trabucco , Albert Qu , Simon Li , Ganeshkumar Ashokavardhanan

Markov reward processes (MRPs) are used to model stochastic phenomena arising in operations research, control engineering, robotics, and artificial intelligence, as well as communication and transportation networks. In many of these cases,…

Machine Learning · Statistics 2020-09-17 Ashwin Pananjady , Martin J. Wainwright

We consider infinite-state turn-based stochastic games of two players, Box and Diamond, who aim at maximizing and minimizing the expected total reward accumulated along a run, respectively. Since the total accumulated reward is unbounded,…

Computer Science and Game Theory · Computer Science 2012-08-09 Tomáš Brázdil , Antonín Kučera , Petr Novotný

Probabilistic control design is founded on the principle that a rational agent attempts to match modelled with an arbitrary desired closed-loop system trajectory density. The framework was originally proposed as a tractable alternative to…

Machine Learning · Computer Science 2023-11-16 Tom Lefebvre

We study some ergodicity property of zero-sum stochastic games with a finite state space and possibly unbounded payoffs. We formulate this property in operator-theoretical terms, involving the solvability of an optimality equation for the…

Optimization and Control · Mathematics 2018-11-15 Antoine Hochart

Simple stochastic games are turn-based 2.5-player games with a reachability objective. The basic question asks whether one player can ensure reaching a given target with at least a given probability. A natural extension is games with a…

Computer Science and Game Theory · Computer Science 2021-02-02 Pranav Ashok , Krishnendu Chatterjee , Jan Kretinsky , Maximilian Weininger , Tobias Winkler

We study the trajectory optimization problem under chance constraints for continuous-time stochastic systems. To address chance constraints imposed on the entire stochastic trajectory, we propose a framework based on the set erosion…

Optimization and Control · Mathematics 2025-04-08 Zishun Liu , Liqian Ma , Yongxin Chen

In this paper, we consider a risk-based optimal investment problem of an insurer in a regime-switching jump diffusion model with noisy memory. Using the model uncertainty modeling, we formulate the investment problem as a zero-sum,…

Portfolio Management · Quantitative Finance 2019-03-25 Rodwell Kufakunesu , Calisto Guambe , Lesedi Mabitsela

We study two-layer neural networks in the mean field limit, where the number of neurons tends to infinity. In this regime, the optimization over the neuron parameters becomes the optimization over the probability measures, and by adding an…

Optimization and Control · Mathematics 2023-08-17 Fan Chen , Zhenjie Ren , Songbo Wang

We consider the problem of minimizing a certainty equivalent of the total or discounted cost over a finite and an infinite time horizon which is generated by a Partially Observable Markov Decision Process (POMDP). The certainty equivalent…

Probability · Mathematics 2021-07-21 Nicole Bäuerle , Ulrich Rieder

This article is related to risk-sensitive nonzero-sum stochastic differential games in the Markovian framework. This game takes into account the attitudes of the players toward risk and the utility is of exponential form. We show the…

Optimization and Control · Mathematics 2014-12-04 Said Hamadène , Rui Mu

Many decision-making processes involve evaluating and then selecting items; examples include scientific peer review, job hiring, school admissions, and investment decisions. The eventual selection is performed by applying rules or…

Computer Science and Game Theory · Computer Science 2025-10-23 Alexander Goldberg , Giulia Fanti , Nihar B. Shah

This paper studies constrained Markov decision processes (CMDPs) with constraints against stochastic thresholds, aiming at safety of reinforcement learning in unknown and uncertain environments. We leverage a Growing-Window estimator…

Machine Learning · Computer Science 2025-12-25 Qian Zuo , Fengxiang He

We study optimal execution in markets with transient price impact in a competitive setting with $N$ traders. Motivated by prior negative results on the existence of pure Nash equilibria, we consider randomized strategies for the traders and…

Trading and Market Microstructure · Quantitative Finance 2026-05-19 Steven Campbell , Marcel Nutz

We study the optimal scheduling problem for a Markovian multiclass queueing network with abandonment in the Halfin--Whitt regime, under the long run average (ergodic) risk sensitive cost criterion. The objective is to prove asymptotic…

Probability · Mathematics 2024-10-23 Sumith Reddy Anugu , Guodong Pang

Game-theoretic concepts have been extensively studied in economics to provide insight into competitive behaviour and strategic decision making. As computing systems increasingly involve concurrently acting autonomous agents, game-theoretic…

Formal Languages and Automata Theory · Computer Science 2022-07-01 Marta Kwiatkowska , Gethin Norman , David Parker , Gabriel Santos , Rui Yan
‹ Prev 1 4 5 6 7 8 10 Next ›