中文
相关论文

相关论文: Smooth approximation of feedback laws for infinite…

200 篇论文

For optimal control of diffusions under several criteria, due to computational or analytical reasons, many studies have a apriori assumed control policies to be Lipschitz or smooth, often with no rigorous analysis on whether this…

最优化与控制 · 数学 2024-05-28 Somnath Pradhan , Serdar Yuksel

This paper considers stochastic convex optimization problems with smooth functional constraints arising in constrained estimation and robust signal recovery. We operate in the high-dimensional and highly-constrained setting, where oracle…

最优化与控制 · 数学 2025-12-16 Vaibhav Rajoriya , Prateek Priyaranjan Pradhan , Ketan Rajawat

Safe Reinforcement Learning from Human Feedback (Safe RLHF) has recently achieved empirical success in developing helpful and harmless large language models by decoupling human preferences regarding helpfulness and harmlessness. Existing…

机器学习 · 计算机科学 2026-04-22 Qiang Liu , Adrienne Kline , Ermin Wei

We develop the dynamic programming approach for a family of infinite horizon boundary control problems with linear state equation and convex cost. We prove that the value function of the problem is the unique regular solution of the…

最优化与控制 · 数学 2008-06-27 Silvia Faggian , Fausto Gozzi

In this paper, for POMDPs, we provide the convergence of a Q learning algorithm for control policies using a finite history of past observations and control actions, and, consequentially, we establish near optimality of such limit Q…

机器学习 · 计算机科学 2022-10-27 Ali Devran Kara , Serdar Yuksel

In this paper, we prove both necessary and sufficient maximum principles for infinite horizon discounted control problems of stochastic Volterra integral equations with finite delay and a convex control domain. The corresponding adjoint…

最优化与控制 · 数学 2023-03-15 Yushi Hamaguchi

We present novel results on the solution of a class of leavable, undiscounted optimal control problems in the minimax sense for nonlinear, continuous-state, discrete-time plants. The problem class includes entry-(exit-)time problems as well…

最优化与控制 · 数学 2018-09-05 Gunther Reissig , Matthias Rungger

This article deals with stochastic processes endowed with the Markov (memoryless) property and evolving over general (uncountable) state spaces. The models further depend on a non-deterministic quantity in the form of a control input, which…

系统与控制 · 计算机科学 2015-09-11 Sofie Haesaert , Robert Babuska , Alessandro Abate

We study the problem of mixed $\mathit{H}_2/\mathit{H}_\infty$ control in the infinite-horizon setting. We identify the optimal causal controller that minimizes the $\mathit{H}_2$ cost of the closed-loop system subject to an…

最优化与控制 · 数学 2024-10-01 Vikrant Malik , Taylan Kargin , Joudi Hajar , Babak Hassibi

We introduce a notion of inexact model of a convex objective function, which allows for errors both in the function and in its gradient. For this situation, a gradient method with an adaptive adjustment of some parameters of the model is…

最优化与控制 · 数学 2021-10-12 Fedor S. Stonyakin

We propose a Bayesian framework for feedback boundary control for hyperbolic balance laws. The method propagates a probability distribution over feedback parameters by using Lyapunov decay estimates as a likelihood. In the linear setting,…

数值分析 · 数学 2026-02-03 Markus Bambach , Shaoshuai Chu , Michael Herty , Yunong Lin

This paper considers the problem of finding near-optimal Markovian randomized (MR) policies for finite-state-action, infinite-horizon, constrained risk-sensitive Markov decision processes (CRSMDPs). Constraints are in the form of standard…

最优化与控制 · 数学 2023-03-14 Uday Kumar M , Sanjay P Bhat , Veeraruna Kavitha , Nandyala Hemachandra

This article presents a dynamic regret analysis for stochastic model predictive control (SMPC) in linear systems with quadratic performance index and additive and multiplicative uncertainties. Under a finite support assumption, the problem…

最优化与控制 · 数学 2025-02-04 Sungho Shin , Sen Na , Mihai Anitescu

This work focuses on numerical solutions of optimal control problems. A time discretization error representation is derived for the approximation of the associated value function. It concerns Symplectic Euler solutions of the Hamiltonian…

最优化与控制 · 数学 2016-02-23 Jesper Karlsson , Stig Larsson , Mattias Sandberg , Anders Szepessy , Raùl Tempone

We describe inexact proximal Newton-like methods for solving degenerate regularized optimization problems and for the broader problem of finding a zero of a generalized equation that is the sum of a continuous map and a maximal monotone…

最优化与控制 · 数学 2026-02-12 Ching-pei Lee , Stephen J. Wright

We investigate a limit value of an optimal control problem when the horizon converges to infinity. For this aim, we suppose suitable nonexpansive-like assumptions which does not imply that the limit is independent of the initial state as it…

最优化与控制 · 数学 2009-10-21 Marc Quincampoix , Jérôme Renault

We prove optimality principles for semicontinuous bounded viscosity solutions of Hamilton-Jacobi-Bellman equations. In particular we provide a representation formula for viscosity supersolutions as value functions of suitable obstacle…

最优化与控制 · 数学 2007-05-23 Annalisa Cesaroni

Feedback controllers for port-Hamiltonian systems reveal an intrinsic inverse optimality property since each passivating state feedback controller is optimal with respect to some specific performance index. Due to the nonlinear…

最优化与控制 · 数学 2020-07-20 Lukas Kölsch , Pol Jané Soneira , Felix Strehle , Sören Hohmann

It is a longstanding unsolved problem to characterize the optimal feedback controls for general linear quadratic optimal control problem of stochastic evolution equation with random coefficients. A solution to this problem is given in [21]…

最优化与控制 · 数学 2022-02-22 Qi Lü , Tianxiao Wang

Towards bridging classical optimal control and online learning, regret minimization has recently been proposed as a control design criterion. This competitive paradigm penalizes the loss relative to the optimal control actions chosen by a…

系统与控制 · 电气工程与系统科学 2023-06-27 Andrea Martin , Luca Furieri , Florian Dörfler , John Lygeros , Giancarlo Ferrari-Trecate