中文
相关论文

相关论文: Mimicking and Conditional Control with Hard Killin…

200 篇论文

Domain adaptation in imitation learning represents an essential step towards improving generalizability. However, even in the restricted setting of third-person imitation where transfer is between isomorphic Markov Decision Processes, there…

机器学习 · 计算机科学 2020-03-02 Aaron Zweig , Joan Bruna

The verification theorem serving as an optimality condition for the optimal control problem, has been expected and studied for a long time. The purpose of this paper is to establish this theorem for control systems governed by stochastic…

最优化与控制 · 数学 2022-09-21 Liangying Chen , Qi Lü

The goal of this paper is to analyze distributional Markov Decision Processes as a class of control problems in which the objective is to learn policies that steer the distribution of a cumulative reward toward a prescribed target law,…

最优化与控制 · 数学 2026-02-09 Nicole Bäuerle , Athanasios Vasileiadis

We study a Markov decision problem in which the state space is the set of finite marked point configurations in the plane, the actions represent thinnings, the reward is proportional to the mark sum which is discounted over time, and the…

概率论 · 数学 2023-09-08 M. N. M. van Lieshout

In this paper, we propose a new policy iteration algorithm to compute the value function and the optimal controls of continuous time stochastic control problems. The algorithm relies on successive approximations using linear-quadratic…

最优化与控制 · 数学 2024-09-09 Dylan Possamaï , Ludovic Tangpi

We prove the stochastic domination for determinantal processes associated with finite rank projection kernels. The result was first proved by Lyons in discrete setting. We avoid the machinery of matroids in order to obtain a proof that…

概率论 · 数学 2020-09-22 Raghavendra Tripathi

When the unconditioned process is a diffusion submitted to a space-dependent killing rate $k(\vec x)$, various conditioning constraints can be imposed for a finite time horizon $T$. We first analyze the conditioned process when one imposes…

统计力学 · 物理学 2022-09-01 Alain Mazzolo , Cécile Monthus

Imitation learning (IL) is notably effective for robotic tasks where directly programming behaviors or defining optimal control costs is challenging. In this work, we address a scenario where the imitator relies solely on observed behavior…

机器学习 · 计算机科学 2024-08-20 Rishabh Agrawal , Nathan Dahlin , Rahul Jain , Ashutosh Nayyar

We consider the problem of computing the set of initial states of a dynamical system such that there exists a control strategy to ensure that the trajectories satisfy a temporal logic specification with probability 1 (almost-surely). We…

系统与控制 · 计算机科学 2015-02-24 Maria Svorenova , Jan Kretinsky , Martin Chmelik , Krishnendu Chatterjee , Ivana Cerna , Calin Belta

Here an original idea is suggested to prove the existence of optimal control for some types of non- linear problems. The obtained results can be considered as individual existence theorems (in some sense).

最优化与控制 · 数学 2007-05-23 A. A. Niftiyev

In this article, we consider the deterministic impulsively controlled system with infinite horizon and several discounted objective functionals. The constructed optimal control problem with functional constraints is reformulated as a Markov…

最优化与控制 · 数学 2026-02-10 A. Piunovskiy

Formally verifying the correctness of mathematical proofs is more accessible than ever, however, the learning curve remains steep for many of the state-of-the-art interactive theorem provers (ITP). Deriving the most appropriate subsequent…

计算机科学中的逻辑 · 计算机科学 2024-11-05 Liao Zhang , David M. Cerna , Cezary Kaliszyk

When applying imitation learning techniques to fit a policy from expert demonstrations, one can take advantage of prior stability/robustness assumptions on the expert's policy and incorporate such control-theoretic prior knowledge…

最优化与控制 · 数学 2021-03-25 Aaron Havens , Bin Hu

Fleming-Viot type particle systems represent a classical way to approximate the distribution of a Markov process with killing, given that it is still alive at a final deterministic time. In this context, each particle evolves independently…

概率论 · 数学 2017-09-21 Bernard Delyon , Frédéric Cérou , Arnaud Guyader , Mathias Rousset

In this paper we study the approximate controllability and existence of optimal control of impulsive fractional semilinear delay differential equations with non-local conditions. We use Sadovskii's fixed point theorem, semigroup theory of…

经典分析与常微分方程 · 数学 2014-02-10 Lakshman Mahto , Syed Abbas

In this paper, we generalise Pontryagin's stochastic maximum principle to controlled McKean-Vlasov equations with anticipating law. The associated new type of delayed backward equations with implicit terminal condition is studied.

最优化与控制 · 数学 2017-07-03 Nacira Agram

Nonlinear constrained optimization problems are encountered in many scientific fields. To utilize the huge calculation power of current computers, many mathematic models are also rebuilt as optimization problems. Most of them have…

最优化与控制 · 数学 2011-10-03 Wei Zhang , Xudong Shi , Liwen Wang

We consider optimal control problems for systems governed by mean-field stochastic differential equations, where the control enters both the drift and the diffusion coefficient. We study the relaxed model, in which admissible controls are…

最优化与控制 · 数学 2017-02-02 Khaled Bahlali , Meriem Mezerdi , Brahim Mezerdi

This paper is concerned with a partially observed hybrid optimal control problem, where continuous dynamics and discrete events coexist and in particular, the continuous dynamics can be observed while the discrete events, described by a…

最优化与控制 · 数学 2023-03-14 Siyu Lv , Jie Xiong , Wen Xu

In this survey we present the near-optimal stochastic control problem according to some recent tools in the literature. In particular, we focus on the approach of a discretization of the noise values instead of the canonical…

概率论 · 数学 2021-06-30 Lourival Lima , Paulo Ruffino , Francys Souza