中文
相关论文

相关论文: Dual dynamic programming for stochastic programs o…

200 篇论文

Noisy observations coupled with nonlinear dynamics pose one of the biggest challenges in robot motion planning. By decomposing nonlinear dynamics into a discrete set of local dynamics models, hybrid dynamics provide a natural way to model…

机器人学 · 计算机科学 2018-10-10 Ajinkya Jain , Scott Niekum

In this paper, we investigate dynamic optimization problems featuring both stochastic control and optimal stopping in a finite time horizon. The paper aims to develop new methodologies, which are significantly different from those of mixed…

投资组合管理 · 定量金融 2014-06-27 Xiongfei Jian , Xun Li , Fahuai Yi

Solving large-scale multistage stochastic programming (MSP) problems poses a significant challenge as commonly used stagewise decomposition algorithms, including stochastic dual dynamic programming (SDDP), face growing time complexity as…

机器学习 · 计算机科学 2025-02-12 Chanyeong Kim , Jongwoong Park , Hyunglip Bae , Woo Chang Kim

We address the stochastic transmission expansion planning (STEP) problem under uncertainty in renewable generation capacity and demand. STEP's objective is to minimize total transmission investment and generation costs. To tackle the…

最优化与控制 · 数学 2026-05-12 Yure Rocha , Teobaldo Bulhões , Anand Subramanian , Joaquim Dias Garcia

A widely used heuristic for solving stochastic optimization problems is to use a deterministic rolling horizon procedure, which has been modified to handle uncertainty (e.g. buffer stocks, schedule slack). This approach has been criticized…

最优化与控制 · 数学 2017-03-16 Raymond T. Perkins , Warren B. Powell

We introduce a novel method for handling endpoint constraints in constrained differential dynamic programming (DDP). Unlike existing approaches, our method guarantees quadratic convergence and is exact, effectively managing rank…

最优化与控制 · 数学 2025-03-07 Maria Parilli , Sergi Martinez , Carlos Mastalli

The main goal of this paper is to apply the machinery of variational analysis and generalized differentiation to study infinite horizon stochastic dynamic programming (DP) with discrete time in the Banach space setting without convexity…

最优化与控制 · 数学 2019-09-04 Boris S. Mordukhovich , Nobusumi Sagara

Constrained Markov Decision Processes (CMDPs) formalize sequential decision-making problems whose objective is to minimize a cost function while satisfying constraints on various cost functions. In this paper, we consider the setting of…

机器学习 · 计算机科学 2020-09-25 Krishna C. Kalagarla , Rahul Jain , Pierluigi Nuzzo

For combinatorial optimization problems, model-based paradigms such as mixed-integer programming (MIP) and constraint programming (CP) aim to decouple modeling and solving a problem: the `holy grail' of declarative problem solving. We…

人工智能 · 计算机科学 2026-03-13 Ryo Kuroiwa , J. Christopher Beck

We consider a dynamic programming (DP) approach to approximately solving an infinite-horizon constrained Markov decision process (CMDP) problem with a fixed initial-state for the expected total discounted-reward criterion with a…

最优化与控制 · 数学 2023-08-08 Hyeong Soo Chang

In [13], an Inexact variant of Stochastic Dual Dynamic Programming (SDDP) called ISDDP was introduced which uses approximate (instead of exact with SDDP) primal dual solutions of the problems solved in the forward and backward passes of the…

最优化与控制 · 数学 2021-04-08 Vincent Guigues , Renato Monteiro , Benar Svaiter

A stochastic program typically involves several parameters, including deterministic first-stage parameters and stochastic second-stage elements that serve as input data. These programs are re-solved whenever any input parameter changes.…

最优化与控制 · 数学 2026-03-16 Chhavi Sharma , Harsha Gangammanavar

We propose two novel numerical schemes for approximate implementation of the dynamic programming~(DP) operation concerned with finite-horizon, optimal control of discrete-time systems with input-affine dynamics. The proposed algorithms…

最优化与控制 · 数学 2022-03-18 M. A. S. Kolarijani , P. Mohajerin Esfahani

Large-scale applications of energy density functional (EDF) methods depend on fast and reliable algorithms to solve the associated non-linear self-consistency problem. When dealing with large single-particle variational spaces, existing…

核理论 · 物理学 2019-06-24 W. Ryssens , M. Bender , P. -H. Heenen

This paper investigates an infinite-horizon linear quadratic stochastic (LQS) optimal control problem for a class of continuous-time stochastic systems. By employing the technique of adaptive dynamic programming (ADP), we propose a novel…

最优化与控制 · 数学 2022-10-11 Heng Zhang

Low-rank methods for semidefinite programming (SDP) have gained a lot of interest recently, especially in machine learning applications. Their analysis often involves determinant-based or Schatten-norm penalties, which are hard to implement…

最优化与控制 · 数学 2021-12-07 Mikhail Krechetov , Jakub Marecek , Yury Maximov , Martin Takac

In this paper, we present Score-life programming, a novel theoretical approach for solving reinforcement learning problems. In contrast with classical dynamic programming-based methods, our method can search over non-stationary policy…

机器学习 · 计算机科学 2023-06-28 Abhinav Muraleedharan

We introduce an algorithm called SQDP (Stochastic Quadratic Dynamic Programming) to solve some multistage stochastic optimization problems having strongly convex recourse functions. The algorithm extends the classical Stochastic Dual…

最优化与控制 · 数学 2026-05-21 Vincent Guigues , Adriana Washington

This paper proposes a machine-learning-based solution approach for solving multi-horizon stochastic programs. The approach embeds a deep learning neural network into a multi-horizon stochastic program to approximate the recourse operational…

最优化与控制 · 数学 2025-12-03 Hongyu Zhang , Gabriele Sormani , Enza Messina , Alan King , Francesca Maggioni

We consider dynamic programming problems with finite, discrete-time horizons and prohibitively high-dimensional, discrete state-spaces for direct computation of the value function from the Bellman equation. For the case that the value…

最优化与控制 · 数学 2020-05-25 Denis Lebedev , Paul Goulart , Kostas Margellos