中文
相关论文

相关论文: Pathwise Relaxed Optimal Control of Rough Differen…

200 篇论文

Optimal control and the associated second-order Hamilton-Jacobi-Bellman (HJB) equation are studied for unbounded stochastic evolution systems in Hilbert spaces. A new notion of viscosity solution, featured by absence of B-continuity, is…

最优化与控制 · 数学 2026-02-10 Shanjian Tang , Jianjun Zhou

Autonomous systems have witnessed a rapid increase in their capabilities, but it remains a challenge for them to perform tasks both effectively and safely. The fact that performance and safety can sometimes be competing objectives renders…

系统与控制 · 电气工程与系统科学 2024-12-04 Hao Wang , Adityaya Dhande , Somil Bansal

In this paper, we explore a new class of stochastic control problems characterized by specific control constraints. Specifically, the admissible controls are subject to the ratcheting constraint, meaning they must be non-decreasing over…

最优化与控制 · 数学 2024-12-17 Mingxin Guo , Zuo Quan Xu

Stochastic optimal control problems governed by delay equations with delay in the control are usually more difficult to study than the the ones when the delay appears only in the state. This is particularly true when we look at the…

概率论 · 数学 2015-06-22 Fausto Gozzi , Federica Masiero

Following the recent resurgence in establishing linear control theoretic benchmarks for reinforcement leaning (RL)-based policy optimization (PO) for complex dynamical systems with continuous state and action spaces, an optimal control…

系统与控制 · 电气工程与系统科学 2023-06-30 Leilei Cui , Lekan Molu

The Hamilton Jacobi Bellman Equation (HJB) provides the globally optimal solution to large classes of control problems. Unfortunately, this generality comes at a price, the calculation of such solutions is typically intractible for systems…

最优化与控制 · 数学 2014-09-23 Matanya B. Horowitz , Anil Damle , Joel W. Burdick

This work is concerned with Hamilton-Jacobi equations of evolution type posed in domains and supplemented with boundary conditions. Hamiltonians are coercive but are neither convex nor quasiconvex. We analyse boundary conditions when…

偏微分方程分析 · 数学 2025-01-01 Nicolas Forcadel , Cyril Imbert , Regis Monneau

The naive application of Reinforcement Learning algorithms to continuous control problems -- such as locomotion and manipulation -- often results in policies which rely on high-amplitude, high-frequency control signals, known colloquially…

机器人学 · 计算机科学 2019-02-14 Steven Bohez , Abbas Abdolmaleki , Michael Neunert , Jonas Buchli , Nicolas Heess , Raia Hadsell

We consider a singular control problem with regime switching that arises in problems of optimal investment decisions of cash-constrained firms. The value function is proved to be the unique viscosity solution of the associated…

计算金融 · 定量金融 2016-10-07 Erwan Pierre , Stéphane Villeneuve , Xavier Warin

Convex Q-learning is a recent approach to reinforcement learning, motivated by the possibility of a firmer theory for convergence, and the possibility of making use of greater a priori knowledge regarding policy or value function structure.…

最优化与控制 · 数学 2022-10-18 Fan Lu , Joel Mathias , Sean Meyn , Karanjit Kalsi

We consider a stochastic optimal control problem where the controller can anticipate the evolution of the driving noise over some dynamically changing time window. The controlled state dynamics are understood as a rough differential…

最优化与控制 · 数学 2025-10-07 Peter Bank , Franziska Bielert

We study the problem of generating control laws for systems with unknown dynamics. Our approach is to represent the controller and the value function with neural networks, and to train them using loss functions adapted from the…

机器人学 · 计算机科学 2023-02-21 Selim Engin , Volkan Isler

We study a class of backward stochastic differential equations (BSDEs) driven by a random measure or, equivalently, by a marked point process. Under appropriate assumptions we prove well-posedness and continuous dependence of the solution…

概率论 · 数学 2012-05-24 Fulvia Confortola , Marco Fuhrman

This paper is concerned with stochastic impulse control problems in which the running cost changes depending on the impulse control. Because of such a dependence, it brings several difficulties when the usual dynamic programming principle…

最优化与控制 · 数学 2025-11-11 Yuchen Cao , Jiongmin Yong

Reward shaping has been applied widely to accelerate Reinforcement Learning (RL) agents' training. However, a principled way of designing effective reward shaping functions, especially for complex continuous control problems, remains…

机器学习 · 计算机科学 2026-02-12 Mateo Juliani , Mingxuan Li , Elias Bareinboim

We propose and analyze a randomization scheme for a general class of impulse control problems. The solution to this randomized problem is characterized as the fixed point of a compound operator which consists of a regularized nonlocal…

最优化与控制 · 数学 2026-05-26 Haoyang Cao , Yuchao Dong , Zhouhao Yang

Solving the Hamilton-Jacobi-Bellman equation is important in many domains including control, robotics and economics. Especially for continuous control, solving this differential equation and its extension the Hamilton-Jacobi-Isaacs…

机器人学 · 计算机科学 2021-10-06 Michael Lutter , Boris Belousov , Shie Mannor , Dieter Fox , Animesh Garg , Jan Peters

A deep learning approach for the approximation of the Hamilton-Jacobi-Bellman partial differential equation (HJB PDE) associated to the Nonlinear Quadratic Regulator (NLQR) problem. A state-dependent Riccati equation control law is first…

最优化与控制 · 数学 2022-07-20 Anastasia Borovykh , Dante Kalise , Alexis Laignelet , Panos Parpas

Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence of closed-form optimal control laws. Recent work~\cite{gelphman2025end} introduced an implicit…

最优化与控制 · 数学 2026-04-28 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung

In this manuscript we consider a class optimal control problem for stochastic differential delay equations. First, we rewrite the problem in a suitable infinite-dimensional Hilbert space. Then, using the dynamic programming approach, we…

最优化与控制 · 数学 2023-02-20 Filippo de Feo , Salvatore Federico , Andrzej Święch