中文
相关论文

相关论文: Mimicking and Conditional Control with Hard Killin…

200 篇论文

Using results from our companion article [arXiv:1112.4824v2] on a Schauder approach to existence of solutions to a degenerate-parabolic partial differential equation, we solve three intertwined problems, motivated by probability theory and…

概率论 · 数学 2016-04-08 Paul M. N. Feehan , Camelia Pop

We aim at studying approximate null-controllability properties of a particular class of piecewise linear Markov processes (Markovian switch systems). The criteria are given in terms of algebraic invariance and are easily computable. We…

最优化与控制 · 数学 2015-07-03 Dan Goreac , Miguel Martinez

A widely used technique for improving policies is success conditioning, in which one collects trajectories, identifies those that achieve a desired outcome, and updates the policy to imitate the actions taken along successful trajectories.…

人工智能 · 计算机科学 2026-01-27 Daniel Russo

Euclidean Markov decision processes are a powerful tool for modeling control problems under uncertainty over continuous domains. Finite state imprecise, Markov decision processes can be used to approximate the behavior of these infinite…

人工智能 · 计算机科学 2020-06-29 Manfred Jaeger , Giorgio Bacci , Giovanni Bacci , Kim Guldstrand Larsen , Peter Gjøl Jensen

We construct a family of self-similar Markov martingales with given marginal distributions. This construction uses the self-similarity and Markov property of a reference process to produce a family of Markov processes that possess the same…

统计理论 · 数学 2015-06-05 Jie Yen Fan , Kais Hamza , Fima Klebaner

It is well-known that well-posedness of a martingale problem in the class of continuous (or r.c.l.l.) solutions enables one to construct the associated transition probability functions. We extend this result to the case when the martingale…

概率论 · 数学 2007-05-23 Abhay G Bhatt , Rajeeva L Karandikar , B V Rao

We develop a martingale approximation approach to studying the limiting behavior of quadratic forms of Markov chains. We use the technique to examine the asymptotic behavior of lag-window estimators in time series and we apply the results…

概率论 · 数学 2011-08-16 Yves F. Atchade , Matias D. Cattaneo

In this paper we study the stochastic control problem of partially observed (multi-dimensional) stochastic system driven by both Brownian motions and fractional Brownian motions. In the absence of the powerful tool of Girsanov…

最优化与控制 · 数学 2023-08-22 Yueyang Zheng , Yaozhong Hu

In this work we formulate and treat an extension of the Imitation from Observations problem. Imitation from Observations is a generalisation of the well-known Imitation Learning problem where state-only demonstrations are considered. In our…

系统与控制 · 电气工程与系统科学 2022-10-11 Tom Lefebvre

We present a method to generate a robot control strategy that maximizes the probability to accomplish a task. The task is given as a Linear Temporal Logic (LTL) formula over a set of properties that can be satisfied at the regions of a…

最优化与控制 · 数学 2015-03-19 Xu Chu Ding , Stephen L. Smith , Calin Belta , Daniela Rus

Decision processes with incomplete state feedback have been traditionally modeled as Partially Observable Markov Decision Processes. In this paper, we present an alternative formulation based on probabilistic regular languages. The proposed…

最优化与控制 · 数学 2009-08-07 Ishanu Chattopadhyay , Asok Ray

We revisit closed-loop performance guarantees for Model Predictive Control in the deterministic and stochastic cases, which extend to novel performance results applicable to receding horizon control of Partially Observable Markov Decision…

最优化与控制 · 数学 2020-05-01 Martin A. Sehr , Robert R. Bitmead

This paper studies the approximation of optimal control policies by quantized (discretized) policies for a very general class of Markov decision processes (MDPs). The problem is motivated by applications in networked control systems,…

最优化与控制 · 数学 2015-05-14 Naci Saldi , Serdar Yüksel , Tamás Linder

A continuous-time Markov process $X$ can be conditioned to be in a given state at a fixed time $T > 0$ using Doob's $h$-transform. This transform requires the typically intractable transition density of $X$. The effect of the $h$-transform…

概率论 · 数学 2024-09-16 Marc Corstanje , Frank van der Meulen , Moritz Schauer

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

机器学习 · 计算机科学 2025-12-24 Peter N. Loxley

We establish a refined version of the Second Law of Thermodynamics for Langevin stochastic processes describing mesoscopic systems driven by conservative or non-conservative forces and interacting with thermal noise. The refinement is based…

Maximum likelihood constraint inference is a powerful technique for identifying unmodeled constraints that affect the behavior of a demonstrator acting under a known objective function. However, it was originally formulated only for…

机器人学 · 计算机科学 2021-09-13 Kaylene C. Stocking , David L. McPherson , Robert P. Matthew , Claire J. Tomlin

In this note, we consider an optimal control problem associated to a differential equation driven by a H\"{o}lder continuous function g of index greater than 1/2. We split our study in two cases. If the coefficient of dg\_t does not depend…

概率论 · 数学 2007-05-23 Laurent Mazliak , Ivan Nourdin

We introduce the use of reinforcement learning for indirect mechanisms, working with the existing class of sequential price mechanisms, which generalizes both serial dictatorship and posted price mechanisms and essentially characterizes all…

计算机科学与博弈论 · 计算机科学 2021-05-07 Gianluca Brero , Alon Eden , Matthias Gerstgrasser , David C. Parkes , Duncan Rheingans-Yoo

The long time behavior of an absorbed Markov process is well described by the limiting distribution of the process conditioned to not be killed when it is observed. Our aim is to give an approximation's method of this limit, when the…

概率论 · 数学 2009-05-25 Denis Villemonais