中文
相关论文

相关论文: Lagrangian Dual Decision Rules for Multistage Stoc…

200 篇论文

The selection of branching variables is a key component of branch-and-bound algorithms for solving Mixed-Integer Programming (MIP) problems since the quality of the selection procedure is likely to have a significant effect on the size of…

最优化与控制 · 数学 2016-08-23 Pierre Le Bodic , George L. Nemhauser

In this paper, we study the learning of safe policies in the setting of reinforcement learning problems. This is, we aim to control a Markov Decision Process (MDP) of which we do not know the transition probabilities, but we have access to…

系统与控制 · 电气工程与系统科学 2022-01-14 Santiago Paternain , Miguel Calvo-Fullana , Luiz F. O. Chamon , Alejandro Ribeiro

We study the minmax optimization problem introduced in [22] for computing policies for batch mode reinforcement learning in a deterministic setting. First, we show that this problem is NP-hard. In the two-stage case, we provide two…

系统与控制 · 计算机科学 2012-10-31 Raphael Fonteneau , Damien Ernst , Bernard Boigelot , Quentin Louveaux

In this paper, we consider stochastic monotone Nash games where each player's strategy set is characterized by possibly a large number of explicit convex constraint inequalities. Notably, the functional constraints of each player may depend…

最优化与控制 · 数学 2023-08-25 Zeinab Alizadeh , Afrooz Jalilzadeh , Farzad Yousefian

In this paper, we propose a novel reinforcement- learning algorithm consisting in a stochastic variance-reduced version of policy gradient for solving Markov Decision Processes (MDPs). Stochastic variance-reduced gradient (SVRG) methods…

机器学习 · 计算机科学 2018-06-15 Matteo Papini , Damiano Binaghi , Giuseppe Canonaco , Matteo Pirotta , Marcello Restelli

In this paper, we address the problem of reconfiguring Earth observation satellite constellation systems through multiple stages. The Multi-stage Constellation Reconfiguration Problem (MCRP) aims to maximize the total observation rewards…

最优化与控制 · 数学 2025-07-22 Hang Woon Lee , David O. Williams Rogers , Brycen D. Pearl , Hao Chen , Koki Ho

"Weakly coupled dynamic program" describes a broad class of stochastic optimization problems in which multiple controlled stochastic processes evolve independently but subject to a set of linking constraints imposed on the controls. One…

最优化与控制 · 数学 2014-05-15 Fan Ye , Helin Zhu , Enlu Zhou

We consider a general class of two-stage distributionally robust optimization (DRO) problems where the ambiguity set is constrained by fixed marginal probability laws that are not necessarily discrete. We derive primal and dual formulations…

最优化与控制 · 数学 2025-10-17 Ariel Neufeld , Qikun Xiang

We introduce a general method for relaxing decision diagrams that allows one to bound job sequencing problems by solving a Lagrangian dual problem on a relaxed diagram. We also provide guidelines for identifying problems for which this…

数据结构与算法 · 计算机科学 2019-08-21 J. N. Hooker

For optimal control problems that involve planning and following a trajectory, two degree of freedom (2DOF) controllers are a ubiquitously used control architecture that decomposes the problem into a trajectory generation layer and a…

最优化与控制 · 数学 2023-11-14 Anusha Srikanthan , Vijay Kumar , Nikolai Matni

We introduce the class of multistage stochastic optimization problems with a random number of stages. For such problems, we show how to write dynamic programming equations and detail the Stochastic Dual Dynamic Programming algorithm to…

最优化与控制 · 数学 2019-07-18 Vincent Guigues

Constraint handling remains a key bottleneck in quantum combinatorial optimization. While slack-variable-based encodings are straightforward, they significantly increase qubit counts and circuit depth, challenging the scalability of quantum…

量子物理 · 物理学 2025-08-12 Monit Sharma , Hoong Chuin Lau

The missing data problem is one of the important issues to address for achieving data quality. While imputation-based methods are designed to achieve data completeness, their efficacy is observed to be diminishing as and when there is…

多智能体系统 · 计算机科学 2026-02-02 Durga Keshav , GVD Praneeth , Chetan Kumar Patruni , Vivek Yelleti , U Sai Ram

In this paper, we present a two-phase augmented Lagrangian method, called QSDPNAL, for solving convex quadratic semidefinite programming (QSDP) problems with constraints consisting of a large number of linear equality, inequality…

最优化与控制 · 数学 2017-01-02 Xudong Li , Defeng Sun , Kim-Chuan Toh

In this paper we apply an augmented Lagrange method to a class of semilinear elliptic optimal control problems with pointwise state constraints. We show strong convergence of subsequences of the primal variables to a local solution of the…

最优化与控制 · 数学 2018-10-25 Veronika Karl , Ira Neitzel , Daniel Wachsmuth

The Double Linear Policy (DLP) framework guarantees a Robust Positive Expectation (RPE) under optimized constant-weight designs or admissible prespecified time-varying policies. However, the sequential optimization of these time-varying…

系统与控制 · 电气工程与系统科学 2026-04-02 Tan Chin Hong , Chung-Han Hsieh

In this paper, we present a stochastic augmented Lagrangian approach on (possibly infinite-dimensional) Riemannian manifolds to solve stochastic optimization problems with a finite number of deterministic constraints.We investigate the…

最优化与控制 · 数学 2025-04-01 Caroline Geiersbach , Tim Suchan , Kathrin Welker

The robust constrained Markov decision process (RCMDP) is a recent task-modelling framework for reinforcement learning that incorporates behavioural constraints and that provides robustness to errors in the transition dynamics model through…

机器学习 · 计算机科学 2024-05-16 David M. Bossens

Multistage Stochastic Programming (MSP) is a class of models for sequential decision-making under uncertainty. MSP problems are known for their computational intractability due to the sequential nature of the decision-making structure and…

最优化与控制 · 数学 2021-02-10 Murwan Siddig , Yongjia Song , Amin Khademi

This paper proposes an algorithm to efficiently solve multistage stochastic programs with block separable recourse where each recourse problem is a multistage stochastic program with stage-wise independent uncertainty. The algorithm first…

最优化与控制 · 数学 2025-07-30 Nicolò Mazzi , Ken Mckinnon , Hongyu Zhang