English
Related papers

Related papers: State-Continuity Approximation of Markov Decision …

200 papers

Extensions of Kemeny's constant, as derived for irreducible finite Markov chains in discrete time, to Markov renewal processes and Markov chains in continuous time are discussed. Three alternative Kemeny's functions and their variants are…

Probability · Mathematics 2018-09-17 Jeffrey J Hunter

We develop an approach to time-consistent risk evaluation of continuous-time processes in Markov systems. Our analysis is based on dual representation of coherent risk measures, differentiability concepts for multivalued mappings, and a…

Optimization and Control · Mathematics 2017-01-31 Darinka Dentcheva , Andrzej Ruszczynski

This paper considers robust Markov decision processes under parametric transition distributions. We assume that the true transition distribution is uniquely specified by some parametric distribution, and explicitly enforce that the…

Optimization and Control · Mathematics 2022-11-24 Ben Black , Trivikram Dokka , Christopher Kirkbride

The paper is concerned with a variant of the continuous-time finite state Markov game of control and stopping where both players can affect transition rates, while only one player can choose a stopping time. We use the dynamic programming…

Optimization and Control · Mathematics 2022-08-09 Yurii Averboukh

The solution convergence of Markov Decision Processes (MDPs) can be accelerated by prioritized sweeping of states ranked by their potential impacts to other states. In this paper, we present new heuristics to speed up the solution…

Artificial Intelligence · Computer Science 2019-01-07 Shoubhik Debnath , Lantao Liu , Gaurav Sukhatme

This paper deals with the general discounted impulse control problem of a piecewise deterministic Markov process. We investigate a new family of epsilon-optimal strategies. The construction of such strategies is explicit and only…

Probability · Mathematics 2016-03-28 Benoîte de Saporta , François Dufour , Alizée Geeraert

We are interested in the connection between a metastable continuous state space Markov process (satisfying e.g. the Langevin or overdamped Langevin equation) and a jump Markov process in a discrete state space. More precisely, we use the…

Probability · Mathematics 2017-02-08 Giacomo Di Gesù , Tony Lelièvre , Dorian Le Peutrec , Boris Nectoux

Markov decision processes (MDPs) with rewards are a widespread and well-studied model for systems that make both probabilistic and nondeterministic choices. A fundamental result about MDPs is that their minimal and maximal expected rewards…

Logic in Computer Science · Computer Science 2024-11-26 Kevin Batz , Benjamin Lucien Kaminski , Christoph Matheja , Tobias Winkler

Integrated task and motion planning has emerged as a challenging problem in sequential decision making, where a robot needs to compute high-level strategy and low-level motion plans for solving complex tasks. While high-level strategies…

Artificial Intelligence · Computer Science 2018-02-19 Siddharth Srivastava , Nishant Desai , Richard Freedman , Shlomo Zilberstein

We present an algorithm that, given a representation of a road network in lane-level detail, computes a route that minimizes the expected cost to reach a given destination. In doing so, our algorithm allows us to solve for the complex…

Robotics · Computer Science 2023-07-14 Mitchell Jones , Maximilian Haas-Heger , Jur van den Berg

In this paper we consider a broad class of infinite horizon discrete-time optimal control models that involve a nonnegative cost function and an affine mapping in their dynamic programming equation. They include as special cases classical…

Optimization and Control · Mathematics 2017-11-29 Dimitri Bertsekas

We propose a new method for simulating electron dynamics in open quantum systems out of equilibrium, using a finite atomistic model. The proposed method is motivated by the intuitive and practical nature of the driven Liouville von-Neumann…

Mesoscale and Nanoscale Physics · Physics 2014-09-23 Tamar Zelovich , Leeor Kronik , Oded Hod

The contribution gives a micro-structural insight into the pedestrian decision process during an egress situation. A method how to extract the decisions of pedestrians from the trajectories recorded during the experiments is introduced. The…

Multiagent Systems · Computer Science 2018-01-08 Pavel Hrabák , Ondřej Ticháček , Vladimíra Sečkárová

This paper deals with the optimal stopping problem under partial observation for piecewise-deterministic Markov processes. We first obtain a recursive formulation of the optimal filter process and derive the dynamic programming equation of…

Probability · Mathematics 2013-05-28 Adrien Brandejsky , Benoîte de Saporta , François Dufour

The paper is concerned with a zero-sum continuous-time stochastic differential game with a dynamics controlled by a Markov process and a terminal payoff. The value function of the original game is estimated using the value function of a…

Optimization and Control · Mathematics 2016-02-16 Yurii Averboukh

We consider the problem of finding an input signal which transfers a linear boundary controlled 1D parabolic partial differential equation with spatially-varying coefficients from a given initial state to a desired final state. The initial…

Systems and Control · Electrical Eng. & Systems 2024-05-17 Soham Chatterjee , Vivek Natarajan

Dynamical system state estimation and parameter calibration problems are ubiquitous across science and engineering. Bayesian approaches to the problem are the gold standard as they allow for the quantification of uncertainties and enable…

Data Analysis, Statistics and Probability · Physics 2024-11-12 Kairui Hao , Ilias Bilionis

We consider the approximation of the performance of random walks in the quarter-plane. The approximation is in terms of a random walk with a product-form stationary distribution, which is obtained by perturbing the transition probabilities…

Probability · Mathematics 2014-09-15 Jasper Goseling , Richard J. Boucherie , Jan-Kees van Ommeren

Inference, prediction and control of complex dynamical systems from time series is important in many areas, including financial markets, power grid management, climate and weather modeling, or molecular dynamics. The analysis of such highly…

Machine Learning · Statistics 2019-08-19 Hao Wu , Frank Noé

In this paper, we propose a class of efficient, accurate, and general methods for solving state-estimation problems with equality and inequality constraints. The methods are based on recent developments in variable splitting and partially…

Optimization and Control · Mathematics 2020-12-02 Rui Gao , Filip Tronarp , Simo Särkkä