English
Related papers

Related papers: Mimicking and Conditional Control with Hard Killin…

200 papers

Using results from our companion article [arXiv:1112.4824v2] on a Schauder approach to existence of solutions to a degenerate-parabolic partial differential equation, we solve three intertwined problems, motivated by probability theory and…

Probability · Mathematics 2016-04-08 Paul M. N. Feehan , Camelia Pop

We aim at studying approximate null-controllability properties of a particular class of piecewise linear Markov processes (Markovian switch systems). The criteria are given in terms of algebraic invariance and are easily computable. We…

Optimization and Control · Mathematics 2015-07-03 Dan Goreac , Miguel Martinez

A widely used technique for improving policies is success conditioning, in which one collects trajectories, identifies those that achieve a desired outcome, and updates the policy to imitate the actions taken along successful trajectories.…

Artificial Intelligence · Computer Science 2026-01-27 Daniel Russo

Euclidean Markov decision processes are a powerful tool for modeling control problems under uncertainty over continuous domains. Finite state imprecise, Markov decision processes can be used to approximate the behavior of these infinite…

Artificial Intelligence · Computer Science 2020-06-29 Manfred Jaeger , Giorgio Bacci , Giovanni Bacci , Kim Guldstrand Larsen , Peter Gjøl Jensen

We construct a family of self-similar Markov martingales with given marginal distributions. This construction uses the self-similarity and Markov property of a reference process to produce a family of Markov processes that possess the same…

Statistics Theory · Mathematics 2015-06-05 Jie Yen Fan , Kais Hamza , Fima Klebaner

It is well-known that well-posedness of a martingale problem in the class of continuous (or r.c.l.l.) solutions enables one to construct the associated transition probability functions. We extend this result to the case when the martingale…

Probability · Mathematics 2007-05-23 Abhay G Bhatt , Rajeeva L Karandikar , B V Rao

We develop a martingale approximation approach to studying the limiting behavior of quadratic forms of Markov chains. We use the technique to examine the asymptotic behavior of lag-window estimators in time series and we apply the results…

Probability · Mathematics 2011-08-16 Yves F. Atchade , Matias D. Cattaneo

In this paper we study the stochastic control problem of partially observed (multi-dimensional) stochastic system driven by both Brownian motions and fractional Brownian motions. In the absence of the powerful tool of Girsanov…

Optimization and Control · Mathematics 2023-08-22 Yueyang Zheng , Yaozhong Hu

In this work we formulate and treat an extension of the Imitation from Observations problem. Imitation from Observations is a generalisation of the well-known Imitation Learning problem where state-only demonstrations are considered. In our…

Systems and Control · Electrical Eng. & Systems 2022-10-11 Tom Lefebvre

We present a method to generate a robot control strategy that maximizes the probability to accomplish a task. The task is given as a Linear Temporal Logic (LTL) formula over a set of properties that can be satisfied at the regions of a…

Optimization and Control · Mathematics 2015-03-19 Xu Chu Ding , Stephen L. Smith , Calin Belta , Daniela Rus

Decision processes with incomplete state feedback have been traditionally modeled as Partially Observable Markov Decision Processes. In this paper, we present an alternative formulation based on probabilistic regular languages. The proposed…

Optimization and Control · Mathematics 2009-08-07 Ishanu Chattopadhyay , Asok Ray

We revisit closed-loop performance guarantees for Model Predictive Control in the deterministic and stochastic cases, which extend to novel performance results applicable to receding horizon control of Partially Observable Markov Decision…

Optimization and Control · Mathematics 2020-05-01 Martin A. Sehr , Robert R. Bitmead

This paper studies the approximation of optimal control policies by quantized (discretized) policies for a very general class of Markov decision processes (MDPs). The problem is motivated by applications in networked control systems,…

Optimization and Control · Mathematics 2015-05-14 Naci Saldi , Serdar Yüksel , Tamás Linder

A continuous-time Markov process $X$ can be conditioned to be in a given state at a fixed time $T > 0$ using Doob's $h$-transform. This transform requires the typically intractable transition density of $X$. The effect of the $h$-transform…

Probability · Mathematics 2024-09-16 Marc Corstanje , Frank van der Meulen , Moritz Schauer

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

Machine Learning · Computer Science 2025-12-24 Peter N. Loxley

We establish a refined version of the Second Law of Thermodynamics for Langevin stochastic processes describing mesoscopic systems driven by conservative or non-conservative forces and interacting with thermal noise. The refinement is based…

Maximum likelihood constraint inference is a powerful technique for identifying unmodeled constraints that affect the behavior of a demonstrator acting under a known objective function. However, it was originally formulated only for…

Robotics · Computer Science 2021-09-13 Kaylene C. Stocking , David L. McPherson , Robert P. Matthew , Claire J. Tomlin

In this note, we consider an optimal control problem associated to a differential equation driven by a H\"{o}lder continuous function g of index greater than 1/2. We split our study in two cases. If the coefficient of dg\_t does not depend…

Probability · Mathematics 2007-05-23 Laurent Mazliak , Ivan Nourdin

We introduce the use of reinforcement learning for indirect mechanisms, working with the existing class of sequential price mechanisms, which generalizes both serial dictatorship and posted price mechanisms and essentially characterizes all…

Computer Science and Game Theory · Computer Science 2021-05-07 Gianluca Brero , Alon Eden , Matthias Gerstgrasser , David C. Parkes , Duncan Rheingans-Yoo

The long time behavior of an absorbed Markov process is well described by the limiting distribution of the process conditioned to not be killed when it is observed. Our aim is to give an approximation's method of this limit, when the…

Probability · Mathematics 2009-05-25 Denis Villemonais
‹ Prev 1 4 5 6 7 8 10 Next ›