中文
相关论文

相关论文: Mimicking and Conditional Control with Hard Killin…

200 篇论文

The main goal of this paper is to apply the so-called policy iteration algorithm (PIA) for the long run average continuous control problem of piecewise deterministic Markov processes (PDMP's) taking values in a general Borel space and with…

概率论 · 数学 2009-02-17 O. L. V. Costa , F. Dufour

Stochastic optimal control of dynamical systems is a crucial challenge in sequential decision-making. Recently, control-as-inference approaches have had considerable success, providing a viable risk-sensitive framework to address the…

机器学习 · 计算机科学 2023-12-22 Hany Abdulsamad , Sahel Iqbal , Adrien Corenflos , Simo Särkkä

This article introduces an imitation learning method for learning maximum entropy policies that comply with constraints demonstrated by expert trajectories executing a task. The formulation of the method takes advantage of results…

机器学习 · 计算机科学 2025-07-10 George Papadopoulos , George A. Vouros

In this paper we are concerned with the approximate controllability of a multidimensional semilinear reaction-diffusion equation governed by a multiplicative control, which is locally distributed in the reaction term. For a given initial…

最优化与控制 · 数学 2020-06-26 Mohamed Ouzahra

Complex systems may often be characterized by their hierarchical dynamics. In this paper do we present a method and an operational algorithm that automatically infer this property in a broad range of systems; discrete stochastic processes.…

适应与自组织系统 · 物理学 2007-05-23 Olof Görnerup , Martin Nilsson Jacobi

We consider a hidden Markov model with multiple observation processes, one of which is chosen at each point in time by a policy---a deterministic function of the information state---and attempt to determine which policy minimises the…

概率论 · 数学 2015-03-17 James Y. Zhao

Modeling the purposeful behavior of imperfect agents from a small number of observations is a challenging task. When restricted to the single-agent decision-theoretic setting, inverse optimal control techniques assume that observed behavior…

计算机科学与博弈论 · 计算机科学 2013-08-19 Kevin Waugh , Brian D. Ziebart , J. Andrew Bagnell

Multi-objective optimization models that encode ordered sequential constraints provide a solution to model various challenging problems including encoding preferences, modeling a curriculum, and enforcing measures of safety. A recently…

人工智能 · 计算机科学 2022-09-16 Kyle Hollins Wray , Stas Tiomkin , Mykel J. Kochenderfer , Pieter Abbeel

This paper deals with partially-observed optimal control problems for the state governed by stochastic differential equation with delay. We develop a stochastic maximum principle for this kind of optimal control problems using a variational…

最优化与控制 · 数学 2020-10-15 Shuaiqi Zhang , Xun Li , Jie Xiong

In distributed model predictive control (DMPC), where a centralized optimization problem is solved in distributed fashion using dual decomposition, it is important to keep the number of iterations in the solution algorithm, i.e. the amount…

最优化与控制 · 数学 2013-07-11 Pontus Giselsson , Anders Rantzer

The paper puts forward sufficient conditions for local controllability of a control dynamical system. The results obtained are meaningful in the case when the linear approximation to this system is not completely controllable. As a…

最优化与控制 · 数学 2017-09-05 E. R. Avakov , G. G. Magaril-Il'yaev

We consider optimal control problems, where the control appears in the main part of the operator. We derive the Pontryagin maximum principle as a necessary optimality condition. The proof uses the concept of topological derivatives. In…

最优化与控制 · 数学 2024-08-01 Daniel Wachsmuth

We propose to synthesize a control policy for a Markov decision process (MDP) such that the resulting traces of the MDP satisfy a linear temporal logic (LTL) property. We construct a product MDP that incorporates a deterministic Rabin…

系统与控制 · 计算机科学 2014-09-22 Dorsa Sadigh , Eric S. Kim , Samuel Coogan , S. Shankar Sastry , Sanjit A. Seshia

In this paper we prove a weak necessary and sufficient maximum principle for Markovian regime switching stochastic optimal control problems. Instead of insisting on the maximum condition of the Hamiltonian, we show that 0 belongs to the sum…

最优化与控制 · 数学 2013-09-17 Yusong Li , Harry Zheng

This paper deals with control of partially observable discrete-time stochastic systems. It introduces and studies Markov Decision Processes with Incomplete Information and with semi-uniform Feller transition probabilities. The important…

最优化与控制 · 数学 2022-08-30 Eugene A. Feinberg , Pavlo O. Kasyanov , Michael Z. Zgurovsky

We present the observation that the process of stochastic model predictive control can be formulated in the framework of iterated function systems. The latter has a rich ergodic theory that can be applied to study the system's long-run…

最优化与控制 · 数学 2022-10-14 Vyacheslav Kungurtsev , Jakub Marecek , Robert Shorten

The article poses a general model for optimal control subject to information constraints, motivated in part by recent work of Sims and others on information-constrained decision-making by economic agents. In the average-cost optimal control…

最优化与控制 · 数学 2016-02-24 Ehsan Shafieepoorfard , Maxim Raginsky , Sean P. Meyn

This paper deals with the optimal control of systems governed by nonlinear systems of conservation laws at junctions. The applications considered range from gas compressors in pipelines to open channels management. The existence of an…

偏微分方程分析 · 数学 2008-02-26 R. M. Colombo , G. Guerra , M. Herty , V. Sachers

This paper studies the dynamic programming principle using the measurable selection method for stochastic control of continuous processes. The novelty of this work is to incorporate intermediate expectation constraints on the canonical…

最优化与控制 · 数学 2020-04-22 Yuk-Loong Chow , Xiang Yu , Chao Zhou

Inspired in our work on the controllability for the semilinear with memory \cite{Carrasco-Guevara-Leiva:2017aa, Guevara-Leiva:2016aa, Guevara-Leiva:2017aa}, we present the general cases for the approximate controllability of impulsive…

最优化与控制 · 数学 2020-07-14 Cristi D. Guevara , Hugo Leiva