English
Related papers

Related papers: The Value Functions of Markov Decision Processes

200 papers

This paper presents a way of solving Markov Decision Processes that combines state abstraction and temporal abstraction. Specifically, we combine state aggregation with the options framework and demonstrate that they work well together and…

Artificial Intelligence · Computer Science 2015-01-19 Kamil Ciosek , David Silver

In multiple criteria decision aiding, very often the alternatives are compared by means of a value function compatible with the preferences expressed by the Decision Maker. The problem is that, in general, there is a plurality of compatible…

Optimization and Control · Mathematics 2023-04-14 Sally Giuseppe Arcidiacono , Salvatore Corrente , Salvatore Greco

We prove several results concerning classifications, based on successive observations $(X_1,..., X_n)$ of an unknown stationary and ergodic process, for membership in a given class of processes, such as the class of all finite order Markov…

Probability · Mathematics 2008-06-19 Gusztav Morvai , Benjamin Weiss

In this short article we present some properties regarding the order and the type of an entire function.

Complex Variables · Mathematics 2021-09-07 Vassilis G. Papanicolaou , Eva Kallitsi , George Smyrlis

We provide a characterization of the set of real-valued functions that can be the value function of some polynomial game. Specifically, we prove that a function $u : \dR \to \dR$ is the value function of some polynomial game if and only if…

Optimization and Control · Mathematics 2019-10-16 Galit Ashkenazi-Golan , Eilon Solan , Anna Zseleva

This paper describes the structure of optimal policies for infinite-state Markov Decision Processes with setwise continuous transition probabilities. The action sets may be noncompact. The objective criteria are either the expected total…

Optimization and Control · Mathematics 2021-08-03 Eugene A. Feinberg , Pavlo O. Kasyanov

We give a concise self-contained presentation of known and new limit theorems for the one-type Markov branching processes with continuous time. The new streamlined proofs are based on what we call, the tail generating function approach. Our…

Probability · Mathematics 2014-10-07 Serik Sagitov

Markov decision processes (MDPs) are a standard model for sequential decision-making problems and are widely used across many scientific areas, including formal methods and artificial intelligence (AI). MDPs do, however, come with the…

Artificial Intelligence · Computer Science 2024-12-11 Marnix Suilen , Thom Badings , Eline M. Bovy , David Parker , Nils Jansen

Recently, a class of stochastic processes known as piecewise deterministic Markov processes has been used to define continuous-time Markov chain Monte Carlo algorithms with a number of attractive properties, including compatibility with…

Computation · Statistics 2019-06-03 Alexander Terenin , Daniel Thorngren

Dynamic programming is a class of algorithms used to compute optimal control policies for Markov decision processes. Dynamic programming is ubiquitous in control theory, and is also the foundation of reinforcement learning. In this paper,…

Category Theory · Mathematics 2023-08-01 Jules Hedges , Riu Rodríguez Sakamoto

We consider the down/up crossing property of weighted Markov branching processes. The joint probability distribution of multi crossing numbers of such processes are obtained. In particular, for Markov branching processes, the probability…

Probability · Mathematics 2020-04-20 Yanyun Li , Junping Li

Valuation-Based~System can represent knowledge in different domains including probability theory, Dempster-Shafer theory and possibility theory. More recent studies show that the framework of VBS is also appropriate for representing and…

Artificial Intelligence · Computer Science 2019-09-27 Mieczysław A. Kłopotek , Sławomir T. Wierzchoń

In this paper, we provide strong $L_2$-rates of approximation of the integral-type functionals of Markov processes by integral sums. We improve the method developed in [2]. Under assumptions on the process formulated only in terms of its…

Probability · Mathematics 2015-08-13 Iurii Ganychenko

We start from the observation that, anytime two Markov generators share an eigenvalue, the function constructed from the product of the two eigenfunctions associated to this common eigenvalue is a duality function. We push further this…

Probability · Mathematics 2023-09-08 Frank Redig , Federico Sau

The paper is concerned with a zero-sum continuous-time stochastic differential game with a dynamics controlled by a Markov process and a terminal payoff. The value function of the original game is estimated using the value function of a…

Optimization and Control · Mathematics 2016-02-16 Yurii Averboukh

Value-based methods play a fundamental role in Markov decision processes (MDPs) and reinforcement learning (RL). In this paper, we present a unified control-theoretic framework for analyzing valued-based methods such as value computation…

Optimization and Control · Mathematics 2022-02-15 Xingang Guo , Bin Hu

We develop a qualitative theory of Markov Decision Processes (MDPs) and Partially Observable MDPs that can be used to model sequential decision making tasks when only qualitative information is available. Our approach is based upon an…

Artificial Intelligence · Computer Science 2013-01-07 Blai Bonet , Judea Pearl

The construction presented in this paper can be briefly described as follows: starting from any "finite-dimensional" Markov transition function p_t, on a measurable state space (E,B), we construct a strong Markov process on a certain…

Probability · Mathematics 2013-03-13 Robert J. Vanderbei

This research aimed to introduce the concept of harmonically m-concave set-valued functions, which is obtained from the combination of two definitions: harmonically m-concave functions and set-valued functions. In this work some properties…

Functional Analysis · Mathematics 2024-03-13 Gabriel Santana , Maira Valera-López , Nelson Merentes

The Markov assumption (MA) is fundamental to the empirical validity of reinforcement learning. In this paper, we propose a novel Forward-Backward Learning procedure to test MA in sequential decision making. The proposed test does not assume…

Machine Learning · Statistics 2020-02-06 Chengchun Shi , Runzhe Wan , Rui Song , Wenbin Lu , Ling Leng