中文
相关论文

相关论文: IsoCost-Based Dynamic Programming for Solving Infi…

200 篇论文

We consider a discounted infinite horizon optimal stopping problem. If the underlying distribution is known a priori, the solution of this problem is obtained via dynamic programming (DP) and is given by a well known threshold rule. When…

机器学习 · 计算机科学 2021-02-23 Daniel Russo , Assaf Zeevi , Tianyi Zhang

The optimal \(H_{\infty}\) control problem over an infinite time horizon, which incorporates a performance function with a discount factor \(e^{-\alpha t}\) (\(\alpha > 0\)), is important in various fields. Solving this optimal…

最优化与控制 · 数学 2024-10-04 Guoyuan Chen , Yi Wang , Qinglong Zhou

This paper develops algorithms for high-dimensional stochastic control problems based on deep learning and dynamic programming. Unlike classical approximate dynamic programming approaches, we first approximate the optimal policy by means of…

概率论 · 数学 2021-09-21 Côme Huré , Huyên Pham , Achref Bachouch , Nicolas Langrené

This paper proposes a new indirect solution method for solving state-constrained optimal control problems by revisiting the well-established optimal control theory and addressing the long-standing issue of discontinuous control and costate…

最优化与控制 · 数学 2024-03-08 Kenshiro Oguri

This paper studies the infinite-horizon adaptive optimal control of continuous-time linear periodic (CTLP) systems. A novel value iteration (VI) based off-policy ADP algorithm is proposed for a general class of CTLP systems, so that…

系统与控制 · 电气工程与系统科学 2024-12-20 Bo Pang , Zhong-Ping Jiang

In this paper we propose an on-line policy iteration (PI) algorithm for finite-state infinite horizon discounted dynamic programming, whereby the policy improvement operation is done on-line, only for the states that are encountered during…

最优化与控制 · 数学 2021-06-03 Dimitri Bertsekas

We use one-step conditional risk mappings to formulate a risk averse version of a total cost problem on a controlled Markov process in discrete time infinite horizon. The nonnegative one step costs are assumed to be lower semi-continuous…

最优化与控制 · 数学 2018-06-05 Kerem Ugurlu

Policy iteration (PI) is a widely used algorithm for synthesizing optimal feedback control policies across many engineering and scientific applications. When PI is deployed on infinite-horizon, nonlinear, autonomous optimal-control…

最优化与控制 · 数学 2025-07-15 Tobias Ehring , Behzad Azmi , Bernard Haasdonk

We use the continuation and bifurcation package pde2path to numerically analyze infinite time horizon optimal control problems for parabolic systems of PDEs. The basic idea is a two step approach to the canonical systems, derived from…

最优化与控制 · 数学 2019-12-25 Hannes Uecker , Hannes de Witt

This paper studies an optimal stochastic impulse control problem in a finite horizon with a decision lag, by which we mean that after an impulse is made, a fixed number units of time has to be elapsed before the next impulse is allowed to…

最优化与控制 · 数学 2021-02-09 Chang Li , Jiongmin Yong

A linear control system with quadratic cost functional over infinite time horizon is considered without assuming controllability/stabilizability condition and the global integrability condition for the nonhomogeneous term of the state…

最优化与控制 · 数学 2020-08-25 Jianping Huang , Jiongmin Yong , Hua-Cheng Zhou

Recent work [Ran22] formulated a class of optimal control problems involving positive linear systems, linear stage costs, and elementwise constraints on control. It was shown that the problem admits linear optimal cost and the associated…

最优化与控制 · 数学 2023-09-27 Yuchao Li , Anders Rantzer

In robot-assisted minimally invasive surgery (RMIS), inverse kinematics (IK) must satisfy a remote center of motion (RCM) constraint to prevent tissue damage at the incision point. However, most of existing IK methods do not account for the…

机器人学 · 计算机科学 2024-06-17 Jacinto Colan , Ana Davila , Yasuhisa Hasegawa

For combinatorial optimization problems, model-based paradigms such as mixed-integer programming (MIP) and constraint programming (CP) aim to decouple modeling and solving a problem: the `holy grail' of declarative problem solving. We…

人工智能 · 计算机科学 2026-03-13 Ryo Kuroiwa , J. Christopher Beck

This paper details a methodology to transcribe an optimal control problem into a nonlinear program for generation of the trajectories that optimize a given functional by approximating only the highest order derivatives of a given system's…

最优化与控制 · 数学 2025-09-09 Thomas L. Ahrens , Ian M. Down , Manoranjan Majji

Hyperbolic (HB) programming generalizes many popular convex optimization problems, including semidefinite and second-order cone programming. Despite substantial theoretical progress on HB programming, efficient computational tools for…

最优化与控制 · 数学 2026-02-27 Mehdi Karimi , Levent Tuncel

Non-convex optimal control arises from various applications but may contain multiple stationary points. Classical solvers usually perform a ``local'' search near a saddle point or a local minimum, thus rely on good initial guess to reach…

最优化与控制 · 数学 2025-12-02 Ning Du , Yanlin Liu , Lei Zhang , Xiangcheng Zheng

We introduce a numerically stable reformulation of controllability scoring based on a scaled controllability Gramian, which remains reliably computable even for unstable systems. The resulting optimization problems define dynamics-aware…

最优化与控制 · 数学 2026-01-21 Kota Umezu , Kazuhiro Sato

In this paper we investigate infinite horizon optimal control problems for parametrized partial differential equations. We are interested in feedback control via dynamic programming equations which is well-known to suffer from the curse of…

最优化与控制 · 数学 2018-10-02 Alessandro Alla , Bernard Haasdonk , Andreas Schmidt

We consider a problem of optimal control of an infinite horizon system governed by forward-backward stochastic differential equations with delay. Sufficient and necessary maximum principles for optimal control under partial information in…

最优化与控制 · 数学 2013-12-09 Nacira Agram , Bernt Øksendal