English
Related papers

Related papers: Mimicking and Conditional Control with Hard Killin…

200 papers

We develop a central limit theorem (CLT) for a non-parametric estimator of the transition matrices in controlled Markov chains (CMCs) with finite state-action spaces. Our results establish precise conditions on the logging policy under…

Statistics Theory · Mathematics 2026-03-26 Ziwei Su , Imon Banerjee , Diego Klabjan

Poisson processes and one-dimensional Poisson point processes satisfy three main properties: superposition, thinning, and conditioning. The proof of the first two relies on basic estimates involving the Poisson distribution that are also…

Probability · Mathematics 2025-09-01 Nicolas Lanchier

We consider a general method for the approximation of the distribution of a process conditioned to not hit a given set. Existing methods are based on particle system that are failable, in the sense that, in many situations , they are not…

Probability · Mathematics 2016-06-30 William Oçafrain , Denis Villemonais

Given a general It\^o semimartingale, its Markovian projection is an It\^o process, with Markovian differential characteristics, that matches the one-dimensional marginal laws of the original process. We construct Markovian projections for…

Probability · Mathematics 2024-03-26 Martin Larsson , Shukun Long

This paper focuses on optimizing probabilities of events of interest defined over general controlled discrete-time Markov processes. It is shown that the optimization over a wide class of $\omega$-regular properties can be reduced to the…

Probability · Mathematics 2014-07-22 Ilya Tkachev , Alexandru Mereacre , Joost-Pieter Katoen , Alessandro Abate

For a controllable linear time-varying (LTV) pair $(\boldsymbol{A}_t,\boldsymbol{B}_t)$ and $\boldsymbol{Q}_{t}$ positive semidefinite, we derive the Markov kernel for the It\^{o} diffusion…

Optimization and Control · Mathematics 2025-04-23 Alexis M. H. Teter , Wenqing Wang , Sachin Shivakumar , Abhishek Halder

We present an alternative view for the study of optimal control of partially observed Markov Decision Processes (POMDPs). We first revisit the traditional (and by now standard) separated-design method of reducing the problem to fully…

Optimization and Control · Mathematics 2024-12-20 Serdar Yüksel

Bayesian optimization is a methodology to optimize black-box functions. Traditionally, it focuses on the setting where you can arbitrarily query the search space. However, many real-life problems do not offer this flexibility; in…

A task decomposition method for iterative learning model predictive control is presented. We consider a constrained nonlinear dynamical system and assume the availability of state-input pair datasets which solve a task T1. Our objective is…

Systems and Control · Electrical Eng. & Systems 2020-03-13 Charlott Vallon , Francesco Borrelli

This tutorial describes recently developed general optimality conditions for Markov Decision Processes that have significant applications to inventory control. In particular, these conditions imply the validity of optimality equations and…

Optimization and Control · Mathematics 2016-06-06 Eugene A. Feinberg

A new model for controlled sensing for multihypothesis testing is proposed and studied in the sequential setting. This new model, termed {\em controlled Markovian observation} model, exhibits a more complicated memory structure in the…

Optimization and Control · Mathematics 2014-07-01 Sirin Nitinawarat , Venupogal V. Veeravalli

Optimal control synthesis in stochastic systems with respect to quantitative temporal logic constraints can be formulated as linear programming problems. However, centralized synthesis algorithms do not scale to many practical systems. To…

Systems and Control · Computer Science 2015-03-26 Jie Fu , Shuo Han , Ufuk Topcu

A comparison theorem for state-dependent regime-switching diffusion processes is established, which enables us to control pathwisely the evolution of the state-dependent switching component simply by Markov chains. Moreover, a sharp…

Probability · Mathematics 2024-05-08 Jinghai Shao

The formal verification and controller synthesis for Markov decision processes that evolve over uncountable state spaces are computationally hard and thus generally rely on the use of approximations. In this work, we consider the…

Systems and Control · Computer Science 2018-11-28 Sofie Haesaert , Sadegh Soudjani , Alessandro Abate

Prior work has proposed a simple strategy for reinforcement learning (RL): label experience with the outcomes achieved in that experience, and then imitate the relabeled experience. These outcome-conditioned imitation learning methods are…

Machine Learning · Computer Science 2023-02-21 Benjamin Eysenbach , Soumith Udatha , Sergey Levine , Ruslan Salakhutdinov

Output-Feedback Stochastic Model Predictive Control based on Stochastic Optimal Control for nonlinear systems is computationally intractable because of the need to solve a Finite Horizon Stochastic Optimal Control Problem. However, solving…

Optimization and Control · Mathematics 2020-05-01 Martin A. Sehr , Robert R. Bitmead

Imitation learning often assumes that demonstrations are close to optimal according to some fixed, but unknown, cost function. However, according to satisficing theory, humans often choose acceptable behavior based on their personal (and…

Machine Learning · Computer Science 2025-05-27 Rushit N. Shah , Nikolaos Agadakos , Synthia Sasulski , Ali Farajzadeh , Sanjiban Choudhury , Brian Ziebart

We present the conditions under which the time-optimal control problem for a nonlinear non-autonomous linearizable system can be solved by the method of successive approximations, at each step of which a power Markov moment min-problem is…

Optimization and Control · Mathematics 2022-03-17 Katerina V. Sklyar , Svetlana Yu. Ignatovich

This paper is the first part of our series work to establish pointwise second-order necessary conditions for stochastic optimal controls. In this part, both drift and diffusion terms may contain the control variable but the control region…

Optimization and Control · Mathematics 2014-09-10 Haisen Zhang , Xu Zhang

In this paper, we consider the gradual-impulse control problem of continuous-time Markov decision processes, where the system performance is measured by the expectation of the exponential utility of the total cost. We prove, under very…

Optimization and Control · Mathematics 2023-11-16 Xin Guo , Aiko Kurushima , Alexey Piunovskiy , Yi Zhang
‹ Prev 1 3 4 5 6 7 10 Next ›